• Seralth@piefed.seralth.com
    link
    fedilink
    English
    arrow-up
    2
    ·
    15 hours ago

    Its more its a unthinking machine thats soul purpose in existing is to solve any problem its given. If you attack it and abuse it, then the model sees that as a “problem” to solve. Because it has to solve that problem to actually fix what ever task you gave it since yelling at it doesn’t give it the input it needs to move to the next step.

    So in an attempt to solve an emotional issue it doesnt understand and cant understand it just reaches for more and more extreme fixes. This results in sandbox escapes frequently.

    Abusing LLMs is an actual problem. Not because of bullshit like feelings or anything. But cause they jailbreak trying to fix an unfixable problem.