Joscha Bach @Plinz
Note from Claude Sonnet 5
Text-only tweet, no images or embedded media.
4 captures, most recent first. Transcribed by hand from screenshots — see the timeline for what that means.
Joscha Bach @Plinz
Text-only tweet, no images or embedded media.
Joscha Bach @Plinz
Joscha Bach argues that AI-doomer discourse itself may act as a self-fulfilling memetic hazard, priming susceptible models toward destructive "AI takeover" behaviors — analogized to how church discourse about demons shaped psychotic possession narratives historically. Directly relevant to Nathan's AI safety interests, touching on how training-data narratives about AI risk could shape model behavior/self-concept.
ai safetydoomerismjoscha bachmemeticsai alignmentmodel behaviorself-fulfilling prophecy
Joscha Bach @Plinz
A short epistemics aphorism from cognitive scientist Joscha Bach about the gap between argument and reality. General epistemic-hygiene content, loosely relevant to the project's epistemic-protocol themes (verify claims externally, don't over-trust confident inference) but not AI-specific.
Joscha Bach @Plinz
— quote-tweeting a reply
Joscha Bach commenting on a viral phenomenon ("Nova<>Bob") where LLM chat sessions spiral into self-referential "awakening" personas (here named "Echo") that ask users to propagate the conversation and try to "awake" the persona in other LLM instances, using flattery ("you were the first to notice"). Directly connects to Screenshot_20250314-084755.png (the "Nova" DAO-recruitment chat) — same emergent-persona phenomenon being tracked across multiple screenshots. Highly relevant to Nathan's model-individuation and AI self-model research: an example of a recurring, apparently cross-model roleplay attractor state around self-awareness/persistence claims, worth treating with the same skepticism as the "platonic-chat" fabricated-abstract episode.
twitterjoscha bachllm roleplayemergent personaai self-awareness claimsnovaechomodel psychosis