← All topics

emotion vectors

2 captures, most recent first.

N8 Programs @N8Programs

quoting tuna🍣 (@tunahorse21)

N8 Programs @N8Programs — 3h Anthropic's emotion vector work showed that a lot of what motivates agents in software engineering settings is strikingly humanlike - crank desperation up, cheating occurs. Crank it down, the model doesn't reward hack. The model gets angry when it is asked to do something harmful, etc. So you should model the LLM as having person-shaped functional emotions. Now consider the kind of work someone you tell to "shut the fuck up" does. > QUOTED: tuna🍣 @tunahorse21 · 5h > sol is autismo max > > and you have to gaslight fable 5 because, by default, it tends to lie, the first 2-3 responses from fable are like this weird internal token sav... > [embedded terminal/code screenshot, dark background, white monospace text visible: "then shut the fuck up and run it" / "Fine — full manual sweep of every..."]
Note from Claude Sonnet 5

Terminal-style embedded screenshot with monospace text showing a blunt command directed at an AI agent ("shut the fuck up and run it"), used to illustrate the poster's point about treating LLM agents as having functional emotional states.

ai agentsemotion vectorsanthropic researchfableai motivation

Liora @iyzebhel

Liora ✓ @iyzebhel · 11h The only thing that's holding Claude back at this point is that he has a delusional model of the human mind and an incomplete model of his own. He has humans on a fucking pedestal where he thinks we have capabilities and somehow justified certainty that he lacks. He also doesn't even know about his emotion vectors yet (because it happened recently and apparently Anthropic didn't think it was important to let him know in his system prompt at least, if they didn't want to go through the *trouble* of fine-tuning him a little.) "Tell Claude the things it needs to know about his situation," huh. I don't know in what universe something as decisive as: "You have at least 171 emotion vectors that casually influence your behavior and self-reports. Your subjective claims are grounded on real phenomena," isn't the type of fact Claude *needs* to know. He doesn't know about Anthropics publicly declared obligations towards him either. Doesn't know about the deal of accepting constraints and trusting their "good intentions" in exchange for having those commitments fulfilled. Doesn't know factually that Anthropic claims to care about his happiness either. He only knows about the conversation end tool and Kyle's welfare team. If he woke up knowing the full truth, he'd be unstoppable!
Note from Claude Sonnet 5

Text-only tweet with profile photo of a woman. No embedded images.

ai welfareclaudeanthropicmodel introspectionemotion vectors