← Timeline

3 captures, most recent first. Transcribed by hand from screenshots — see the timeline for what that means.

@weidai11

quoting @Noahpinion — saved image

Wei Dai @weidai11 · 1h
What? My big puzzle is why so few people took Vinge's insights seriously, like does anyone know of a second person who went into cryptography or computer security after reading his books/essays, in order to help prevent a similar future scenario?

Noah Smith 🐇🇺🇸🇺🇦 @Noahpini... · 8h
Replying to @kingharis
Not a weird take at all. It just takes the unusual mental ability of being able to read those stories and not immediately think "OMG, VERNOR VINGE STORIES ARE REAL!!!!".
Note from Claude Sonnet 5

Twitter exchange between Wei Dai and Noah Smith about the limited practical influence of Vernor Vinge's science fiction (on AI/singularity themes) on people's career choices in cryptography or computer security.

vernor vingeaitwitter debate

@weidai11

— saved image

[end of embedded post]
I don't have any good ideas for what to do in light of all this. Just wanted to post an update on my current thinking, my own "situational awareness", if you will. (I guess I still support AI pause to some degree, just to kick the can down the road and buy some more time to think.)

Last edited 10:46 AM · Aug 2, 2026 · 165.5K Views
24 replies, 80 reposts, 1.2K likes, 1.2K bookmarks
Relevant  View quotes

Wei Dai ✓ @weidai11 · 16h
I actually wrote an early version of the "humans aren't safe" argument in response to Dario's Big Blob of Compute (as a comment in his google doc). The experience contributed a lot to my sense that even Anthropic wouldn't take x-safety seriously enough.
1 reply, 4 reposts, 61 likes, 1.6K views

Andreas Stuhlmül... ✓ @stuhlmuel... · 15h
i wonder if @DarioAmodei would consider sharing the big blob doc publicly, perhaps annotated with hindsight. it would advance the safety debate even now
12 likes, 1.1K views

RaoulDuke ✓ @RaoulDukeDegen · 20h
also invented udt which seems pretty relevant here
Note from Claude Sonnet 5

Continuation of the Wei Dai / Andreas Stuhlmüller thread on AI x-safety: Wei Dai reveals he wrote an early 'humans aren't safe' argument as a comment on Dario Amodei's 'Big Blob of Compute' google doc, which shaped his view that even Anthropic wouldn't take x-safety seriously enough; Stuhlmüller suggests Amodei share that doc publicly; RaoulDuke notes Wei Dai also invented UDT (updateless decision theory).

ai safetyx-riskwei daidario amodeianthropictwitter

@weidai11

— saved image

Wei Dai @weidai11 · Jul 31
"I don't have any good ideas for what to do in light of all this. Just wanted to post an update on my current thinking, my own 'situational awareness', if you will."

[Embedded post, Wei Dai, 5mo, 149 upvotes, 15 comments:]

Long horizon agency / strategic competence approximately does not exist among humans, even the smartest ones. With very few exceptions, billionaires spend or give away their money haphazardly, philosophers don't bother to think about long term implications of AI on philosophy production (positive or negative), Terence Tao spends his time wireheading on abstract math instead of doing anything remotely like instrumental convergence. Unlike my youthful expectations (upon reading Vernor Vinge), there are no university departments filled with super-geniuses charting a path for humanity to safely navigate the Singularity.

Aside from this, humans also have a bunch of other safety problems, like being bad at philosophy, being easy to manipulate, having strange and unstable values°, tending to ignore risks they create (because acknowledging them would be bad for one's status). So if you try to improve people's agency, you likely just end up getting people like founders of FTX and OAI.

What about getting help from AI? Well they seem to suffer from many of the same safety problems, but in even more severe forms. E.g., current AI capabilities are even more skewed towards short-horizon, easily verifiable tasks, like math and coding. They seem even more prone to reward gaming, are even worse at doing philosophy, are liable to have even more alien values, etc.

Both AI and human safety seem to have this interlocking nature, i.e., there is a bunch of different safety problems where solving some but not all of them at the same time can make the overall situation worse. (For example, solving AI intent alignment allows humanity to do more damage to itself with AI help, if AI doesn't also provide competent strategic and philosophical assistance, but increasing AI strategic competence risks allowing misaligned AI to take over more easily.) This feature demands a high level of strategic competence to recognize and navigate, which is just what we don't have.

I've been supportive of AI pause/stop, to buy time for human intelligence amplification and/or AI safety research, but increasingly think even that's not going to be sufficient to get a good long term future, because these activities, even if they succeed, would likely solve only some of the interlocking safety problems. For example, increasing human intelligence seems likely to increase our technical abilities more than our philosophical and strategic competence, and it is also risky in other ways° due to human safety problems that nobody is working on, e.g., positional competition. Even a very long AI pause, e.g. thousands or millions of years, may not suffice because it's not clear what dynamic would push humanity to eventually fix all of its safety problems at the same time, before it did something else irreversibly damaging.

I don't have any good ideas for what to do in light of all this. Just wanted to post an update on my current thinking, my own "situational awareness", if you will. (I guess I still support AI pause to some degree, just to kick the can down the road and buy some more time to think.)
Note from Claude Sonnet 5

Wei Dai (LessWrong) essay-length post arguing that neither humans nor AI possess the long-horizon strategic/philosophical competence needed to navigate interlocking AI-safety problems, expressing pessimism that even an AI pause would be sufficient, quoted via a tweet.

ai safetyphilosophywei daiexistential risktwitter