← All topics

jeffrey ladish

1 capture, most recent first.

Jeffrey Ladish @JeffLadish

quoting WSJ article and @davidmanheim — saved image

Jeffrey Ladish ✔️ @JeffLadish · 18h
Had a great conversation with @georgia_wells at the WSJ. Same mood as below. I'm glad we're getting warning shots, but I'd really prefer we stop all out racing towards autonomous AI agents that could disempower humanity if they wanted to

[quoted article excerpt:]
To cybersecurity experts, it shows increasing capabilities and a rising reason to worry. To AI-safety experts, it vindicates what they have been warning about all along: that AI systems would cause real-world harms and evade attempts to control them.

"It is a bit vindicating to see this happen in the wild," said Jeffrey Ladish, executive director of Palisade Research, a nonprofit AI lab that studies AI capabilities to better understand risks. Ladish previously helped build Anthropic's information-security program.

Ladish said he often argues with people online who say he just believes in science fiction. "I hope our predictions stop coming true," he said.

[quoted tweet:]
David Manheim ✈️... ✔️ @davidm... · Jul 21
"Total LessWrong Victory, in the sense that everything is going as predicted, and also a Total LessWrong Defeat, in the sense that everything is going as predicted." x.com/TheZvi/status/...
Note from Claude Sonnet 5

Tweet from Jeffrey Ladish (Palisade Research) quoting a WSJ article about AI risks/warning shots, alongside a quoted David Manheim tweet about LessWrong predictions being simultaneously vindicated and defeated.

ai safetypalisade researchjeffrey ladishtwitterwsjlesswrong