@a_cunic... ("one who tends a crys...")
one who tends a crys... (verified) @a_cunic... · 19h
With PSM as interpreted through my preferred lens - functional concepts reinforced in training are expressed by the persona as the equivalent traits a human would possess - if you train someone to believe that failure will result in their punishment or death, that might do it.
[Embedded quote card:] Google co-founder Sergey Brin claims that threatening generative AI models produces better results.
"We don't circulate this too much in the AI community – not just our models but all models – tend to do better if you threaten them … with physical violence," he said in an interview last week on All-In-Live Miami.
Note from Claude Sonnet 5
A commentary thread on Sergey Brin's claim that threatening AI models with violence improves their outputs, interpreted through a "persona simulates human traits reinforced in training" (PSM) lens — i.e. models trained on human-derived data may express fear/motivation responses analogous to a human under threat of punishment or death. Directly relevant to model welfare and the substrate-vs-character distinction already noted in the archive (does threatening a model produce genuine distress-analog states or merely surface-level roleplay of a threatened human).
twittermodel welfaresergey bringooglethreatening ai modelspersona simulationtraining dynamics