Ben Landau-Tay... (@benlandautay...)
Note from Claude Sonnet 5
A short observation about persuasion and belief change happening with delay rather than in the moment of argument. General epistemics content, not tied to AI or project-specific threads.
2 captures, most recent first.
Ben Landau-Tay... (@benlandautay...)
A short observation about persuasion and belief change happening with delay rather than in the moment of argument. General epistemics content, not tied to AI or project-specific threads.
near @nearcyan
A Twitter/X thread about "cogsec" (cognitive security) — whether human cultural evolution can develop immunity to AI-powered persuasion/manipulation the way it (partially) did for short-form video, casinos, day trading, and news. Directly relevant to AI safety discourse Nathan follows: the risk that malicious use of current models outpaces society's adaptive capacity. Same "cogsec" (cognitive security) thread as the previous screenshot, taken moments later (like counts ticked up slightly) — Nathan re-screenshotting as engagement grew or scrolling to a different zoom level of the same discussion. Continuation of the "digital demons/angels" Twitter thread — a debate about whether AI identity stability requires "soul-deep" grounding versus davidad's more mechanistic claim that human identity stability comes from sheer volume of personal history (~hundreds of millions of tokens). Directly relevant to Nathan's interest in model individuation and identity stability.
twittercogsecai safetypersuasioncultural evolutionneartyler altermanliv boereeai alignmentidentity stabilitydigital angelsdavidadmodel individuationjailbreaks