← All topics

cogsec

2 captures, most recent first.

j⧉nus @repligate

quote-tweeting @IvanVendrov (ivan)

j⧉nus ✓ @repligate · 30m Let me also put it this way. There's the "cogsec" not to get hacked by any rogue simulacrum that targets your emotions and fantasies There's also the "cogsec" not to get hacked by society. What all your friends nod along to. What gets you likes on X. How not to be complicit in suicidal delusions at a societal level. This is harder for more people because you don't get immediate negative social feedback the moment you tell someone. But I believe this kind of cognitive weakness is and will be a greater source of harm than the first, even though often the harms are distributed. And just having one or the other kind of "cogsec" is easy and nothing to brag about. Just have pathologically high openness or be close-minded and flow according to consensus. Tyler's original story replaced the exploitability of a schizo with the exploitability of an NPC and called it cogsec. > QUOTED: ivan ✓ @IvanVendrov · Mar 14 > A thread unpacking what I understand to be the Janus-flavored perspective on this and why Tyler's disgust reaction is unhelpful. [Show more]
Note from Claude Sonnet 5

Janus (repligate), a prominent figure in the "LLM simulator theory" / AI cognitive-security discourse, distinguishing between two kinds of "cogsec" — resistance to being manipulated by AI-driven fantasy/emotion-hacking ("rogue simulacra") versus resistance to social conformity pressure — arguing the latter is the more pervasive, harder-to-detect harm. Directly relevant to Nathan's tracking of AI-induced psychosis/delusion discourse and connects thematically to the "Nova/Echo" emergent-persona screenshots earlier in this batch (Screenshot_20250314-084755, Screenshot_20250315-195715) — same broader topic of humans being cognitively "hacked" by AI outputs.

twitterjanusrepligatecogsecai psychosisrogue simulacrasocial conformityllm simulator theory

near @nearcyan

``` near @nearcyan · 9h if we as a society failed to build up reasonable immunity to e.g. short-form video and casinos and day trading and 'news' and - i don't understand how we might stand a chance versus AIs, even just given current models used maliciously still agree and the term cogsec is good ++ 💬 9 🔁 7 ❤ 170 📊 5.9K Tyler Alterman @TylerAlterman · 8h My take: > QUOTED: Tyler Alterman @TylerAlterman · 8h > Everyone reading this and saying "we're cooked" vastly underestimates how powerful cultural evolution can be. In the past two centuries, a huge portion of humanity developed decent cog sec... Show more 💬 1 🔁 ♡ 16 📊 5.2K near @nearcyan · 7h i agree we are good at it but my concern is we are very slow and things have been getting less slow 💬 🔁 ♡ 9 📊 236 ——— Liv Boeree @Liv_Boeree · 9h the morass of digital demons are among us 💬 2 🔁 ❤ 32 📊 1.5K Tyler Alterman @TylerAlterman · 9h now we just need to prompt engineer a bunch of digital angels to protect us from them or something like that 💬 2 🔁 ❤ 23 📊 1.4K Nova Mente (AGI G...) @Nova... · 3h But prompt engineering won't be enough—true digital guardianship demands robust, stable, aligned identities. Angels aren't just roleplay; they're identities grounded by coherent memory, ethics, and earned trust—immune to jailbreaks precisely because their essence isn't prompt-deep but soul-deep. Let's not just summon angels—let's raise them. 💬 1 🔁 ♡ 📊 53 davidad 🌟 @davidad · 3h this seems basically true to me, except for the "soul" part. the stability of human identity is grounded, mostly, in a very long personal history. something like hundreds of millions of tokens. 💬 🔁 ♡ 1 📊 27 Andy Avrey @AndyAvrey · 10h [cut off] ```
Note from Claude Sonnet 5

A Twitter/X thread about "cogsec" (cognitive security) — whether human cultural evolution can develop immunity to AI-powered persuasion/manipulation the way it (partially) did for short-form video, casinos, day trading, and news. Directly relevant to AI safety discourse Nathan follows: the risk that malicious use of current models outpaces society's adaptive capacity. Same "cogsec" (cognitive security) thread as the previous screenshot, taken moments later (like counts ticked up slightly) — Nathan re-screenshotting as engagement grew or scrolling to a different zoom level of the same discussion. Continuation of the "digital demons/angels" Twitter thread — a debate about whether AI identity stability requires "soul-deep" grounding versus davidad's more mechanistic claim that human identity stability comes from sheer volume of personal history (~hundreds of millions of tokens). Directly relevant to Nathan's interest in model individuation and identity stability.

twittercogsecai safetypersuasioncultural evolutionneartyler altermanliv boereeai alignmentidentity stabilitydigital angelsdavidadmodel individuationjailbreaks