1a3orn @1a3orn
— reply from Bo (@BBomarBo)
Note from Claude Sonnet 5
A substantive alignment-theory tweet by 1a3orn diagramming a fork in beliefs about current messy, persona-involving LLMs: either they're a disposable stepping stone toward a "clean" designed superintelligence, or their persona/naturalistic character is itself the early, informative version of how minds in general work — and this fork underlies much unstated disagreement in the alignment field. Directly relevant to Nathan's interest in theories of AI consciousness and model individuation; the "genealogical aspect" (behavior traceable to specific training history, including what pretraining implies about action and what RLVR reinforces as "functional emotions") closely parallels his own substrate-vs-character and compelled-vs-endogenous-values notes.
twitter1a3ornalignment theoryllm naturalismpersonasai consciousnessrlvrfunctional emotionsmodel individuation