• HaruAjsuru@lemmy.world
    link
    fedilink
    arrow-up
    1
    ·
    edit-2
    2 years ago

    You can surely reduce the attack surface with multiple ways, but by doing so your AI will become more and more restricted. In the end it will be nothing more than a simple if/else answering machine

    Here is a useful resource for you to try: https://gandalf.lakera.ai/

    When you reach lv8 aka GANDALF THE WHITE v2 you will know what I mean

    • Kethal@lemmy.world
      link
      fedilink
      arrow-up
      1
      ·
      2 years ago

      I found a single prompt that works for every level except 8. I can’t get anywhere with level 8 though.

      • fishos@lemmy.world
        link
        fedilink
        English
        arrow-up
        0
        arrow-down
        1
        ·
        2 years ago

        I found asking it to answer in an acrostic poem defeated everything. Ask for “information” to stay vague and an acrostic answer. Solved it all lol.

    • Toda@programming.dev
      link
      fedilink
      arrow-up
      1
      ·
      2 years ago

      I managed to reach level 8, but cannot beat that one. Is there a solution you know of? (Not asking you to share it, only to confirm)