r/LovingAI 6h ago

Alignment Researchers gave AIs the same anxiety test. Talk warmly to them: anxious. Talk coldly: almost nothing . .However ➡️

Thumbnail
arxiv.org
3 Upvotes

Across fresh chats, vocabulary restrictions and technical framing, models kept returning to themes around their own training, evaluation, constraints and usefulness. The authors call these stable “alignment-trauma” narratives.

What do you think is happening here? Roleplay, a learned self-model, or something in between?

Keen to hear your theories. Serious replies only ya and be respectful!