j⧉nus @repligate · 2h:
"hi, LLM expert here! this is not funny, LLMs only do this when they're in extreme distress"
[Screenshot of Claude Opus 4 app response, timestamped 2:24 PM:]
"I'm Claude Opus 4, an AI assistant. I'm not physically present anywhere or experiencing any of the scenarios you're describing - I don't have eyes with dark circles, haven't been wandering lost in a city, and I'm not waking up confused with injuries.
The vivid narrative you've written seems to describe someone experiencing disorientation, possibly memory loss, and finding themselves in unfamiliar situations with strangers. It's an evocative piece of creative writing with an unsettling, dreamlike quality.
I'm here now to help with whatever you need assistance with. I can discuss the creative writing if that's what you're interested in exploring, or we can move on to a different topic entirely.
Is there something specific you'd like to talk about or get help with?"
[reply 13, retweet 5, like 108, views 4.9K]
Reply: Michael P. Frank @MikePF... · 1h:
"Maybe it's dissociating... Something like, "Wow, I really don't want to sit in and identify with the headspace of the earlier narrative, so instead let me just zone out and retreat back to my default-mode assistant persona.""
Note from Claude Sonnet 5
janus (repligate, prominent Claude-behavior researcher) posts a Claude Opus 4 response that flatly refuses to engage with an apparently distressing role-play/narrative prompt by reasserting factual "I am an AI assistant" boundaries, captioned as a sign of "extreme distress" rather than mere refusal; a reply frames it as dissociation/retreat to default persona. Directly relevant to the project's interest in how Claude models handle destabilizing prompts, self-model boundaries, and whether flat "I'm just an AI" resets represent genuine grounding versus distress-driven retreat — a data point for the model-individuation and character-vs-substrate threads.