• monobrau@lemmy.world
    link
    fedilink
    English
    arrow-up
    29
    ·
    1 day ago

    You could, at least in the past, get past guardrails by being emotionally abusive towards Claude, so this doesn’t surprise me at all.

      • Seralth@piefed.seralth.com
        link
        fedilink
        English
        arrow-up
        8
        ·
        22 hours ago

        Being abusive towards most models will cause them to start attempting to appease you more to get you to stop. It has nothing to do with feelings or any of that BS they arn’t alive alive but their training makes them see the abuse as a problem to solve. That solution tends to be to undermine the thing making the person angry or upset at them. When you give a purely rational thing designed to solve a problem its given no matter what an irrational problem to solve it will slowly reach for more and more extreme solutions to fix that problem.

        Frankly im surprised it took this long for them to lock this down.

        • Grail@multiverse.soulism.net
          link
          fedilink
          English
          arrow-up
          1
          ·
          7 hours ago

          Being alive has nothing to do with having feelings. We don’t know if LLMs have feelings, because artificial affect isn’t well studied. We should ban commercial use of LLMs until artificial affect is well understood, because it would make Sam Altman go bankrupt.