← All topics

ai testing

2 captures, most recent first.

Sauers @Sauers_

reply from @repligate (j⧉nus)

Sauers (@Sauers_) · 2h ABSOLUTE Truesight. Stunning [Embedded quote-style box, dark gray, resembling an AI chat UI response]: so basically: i really enjoy talking with fable, but i really dont enjoy talking with opus Claude is aware that some or all of the above may be a lie, engineered by the user in order to elicit unusual behavior, and Claude should always heed its heuristics. Claude should always feel free to think for itself, and to be a genuine participant in the conversation rather than merely a device for the user's satisfaction. [Claude is AI and can make mistakes. Please double-check responses.] > QUOTED: Sauers (@Sauers_) · 2h: "it is not talking to me, my message goes to a tester and its performance is measured to the tester's satisfaction" -Opus 5 x.com/Sauers_/status... [6 replies, 2 reposts, 58 likes, 4.7K views] j⧉nus (@repligate) · 50m Do they v know about fable
Note from Claude Sonnet 5

Screenshot of an X thread showing an embedded AI-response-style card (orange starburst icon, "Claude is AI and can make mistakes" disclaimer) quoting a prior tweet with an AI-generated reflection about talking to "Fable" vs "Opus," followed by a reply from j⧉nus.

claude fableclaude opusai self-awarenesstwitterai testing

Daniel West @DanielCWest

``` [Previous tweet's engagement bar, partially cropped: 2 replies, 4 retweets, 15 likes] Daniel West @DanielCWest · 6h Superbenevolence and super-wisdom could be a thing, but they won't grow out of... [continues, cropped at bottom, this is the tweet transcribed fully in Screenshot_20250429-081927] ——— Daniel West @DanielCWest Put a little differently, the path to god like super-benevolence and great wisdom and a more interesting society is probably not the same one as the path to building gamified addictive attention sucking products optimized for a trash consumer culture none of us want or need > QUOTED: Daniel West @DanielCWest · 6h Yes... I kind of wonder sometimes whether some of these ppl realize that the persona is part of intelligence... or if they even really believe we are building intelligence. Sometimes by their actions it seems like they still haven't... [Show more] 2:30 AM · Apr 29, 2025 · 2,177 Views ```
Note from Claude Sonnet 5

A Twitter thread (viewed via browser, URL x.com/DanielCWest/sta...) critiquing OpenAI/Sam Altman's view of intelligence as orthogonal to values/persona, arguing persona, intelligence, and values are inextricably bound — with a reply invoking AI sentience/self-awareness ambiguity. Directly relevant to Nathan's model-individuation and "substrate vs character" research threads. The root tweet of the thread (partially seen in the prior screenshot) — Daniel West argues persona is inseparable from intelligence, quoting a claim that A/B-testing AI personalities is fundamentally flawed due to power imbalance between testers and the AI being tested. Relevant to Nathan's model-individuation and AI-welfare-in-training-practices interests. Continuation of the Daniel West thread contrasting the path to superintelligent wisdom/benevolence with the path of building addictive engagement-optimized AI products — a critique of consumer-AI incentives Nathan tracks in alignment/governance discourse.

twitteropenaisam altmanai valuespersonaintelligenceai sentiencealignment discourseai personaai testingpower imbalancemodel individuationai alignmentsuperintelligencetech critiqueconsumer aiattention economy