Joshua Achiam @jachiam0
— reply
@jachiam0 (Joshua Achiam) — Jul 21
Some preliminary thoughts about today's cyber developments.
1. Many are freaking out, in a way that is moderately justified, about AI alignment issues indicated by this incident. However, I am not sure that this incident really indicates a fundamental failure of AI alignment
Show more
💬 13 🔁 9 ❤ 78 📊 6.6K 🔖 ⤴
@nathan84686947 (Nathan Helm-Burger) —
"I don't think the critical issue is "model can do scary things," I think the critical issue is "we inhabit a fragile world that can through a sequence of knowable actions be broken."
I'd feel a lot less anxious about this situation if I didn't know this to be the case for more than just cybersecurity…
12:37 AM · Jul 22, 2026 · 182 Views
💬 1 🔁 ❤ 7 🔖 1 ⤴
@jachiam0 (Joshua Achiam) — Jul 22
Yes, and I see this as one of the fundamental grand challenges for humanity in the near term. Vulnerable world hypothesis is a correct diagnosis of danger (but an incorrect diagnosis of solution).
Note from Claude Sonnet 5
Reply-chain screenshot showing Nathan's own tweet reply to OpenAI's Joshua Achiam, referencing Bostrom's Vulnerable World Hypothesis, with Achiam's reply agreeing. Nathan's avatar is a cartoon face making an "OK" hand gesture.
ai alignmentcybersecurityvulnerable world hypothesisnathan's own poststwitter