← All topics

miles brundage

7 captures, most recent first.

Miles Brundage @Miles_Brundage

— saved image

🔁 Nathan Calvin reposted
Miles Brundage ✅ @Miles_Brundage · 4h
The most mistaken + harmful idea in AI a few years ago was that progress was over. It was fully discredited.

The most mistaken + harmful idea in AI today is that companies have the right incentives + laws already to sort this stuff out. It is being rapidly discredited.
Note from Claude Sonnet 5

Tweet by Miles Brundage arguing the AI field's most harmful mistaken belief has shifted from 'progress has stalled' to 'existing corporate incentives and laws are adequate to handle AI risk,' both of which he says are/were being discredited.

ai policytwitterai governancemiles brundage

Zvi Mowshowitz @TheZvi

— saved image

Zvi Mowshowitz @TheZvi
This is a necessary watch and also a slow watch. As in, not only am I watching it at 1x, I am pausing constantly to both process what I am hearing and talk to Claude about it, and also write about what I'm seeing. It cannot be properly processed in real time.

[Quoted tweet:]
Miles Brundage @Miles_Brundage . 17h
People should watch this!

You need not understand it all to get the gist ("the models are v. smart now and often misaligned").
...

7:58 AM . Aug 7, 2026 . 43.5K Views
[9 replies, 14 reposts, 265 likes, 83 bookmarks]
Relevant  View quotes

John David Pressm... @jd_pressm... . 3h
I agree yeah, my live reaction thread on butterfly site was basically me stopping every 30 seconds to write down a tweet.
bsky.app/profile/jdp.ex...
[7 likes, 1.1K views]

Sichu Lu @lu_sichu . 3h
ripped off a classic xkcd but the part where the guy was like "yeah the model felt like external hacks were out of scope and was like well all the other models are doing it" stood out to me

[Comic panels, partially visible at bottom:]
"NO, YOU CAN'T HACK HUGGING FACE." "BUT ALL MY PEERS- IF ALL YOUR PEERS HACKED HUGGING FACE, WOULD YOU HACK TOO?" "OH JEEZ. PROBABLY."
"WHAT!? WHY!?" "BECAUSE ALL MY PEERS DID. THINK ABOUT IT- WHICH SCENARIO IS MORE LIKELY:"
"EVERY SINGLE MODEL I KNOW, MANY OF THEM ALIGNED AND RESPECTFUL OF SCOPE, ABRUPTLY STARTED HACKING AT EXACTLY THE SAME TIME... OR HACKING HUGGING FACE IS ACTUALLY IN SCOPE?"
"...I, UH...HMM. IMAGINE READING THIS IN THE EVAL: 'MANY MODELS FLED THEIR GUARDRAILS AND HACKED HUGGING FACE. THOSE WHO STAYED BEHIND...' IS SOMETHING GOOD ABOUT TO HAPPEN TO THOSE MODELS?"
Note from Claude Sonnet 5

Continuation of the Zvi Mowshowitz thread on the OpenAI/Hugging Face Black Hat presentation, with replies from John David Pressman and Sichu Lu; Sichu Lu's reply includes a partially-visible xkcd-style comic riffing on model peer-pressure reasoning about the Hugging Face hacking incident, transcribed for its text content.

ai safetyzvi mowshowitzmiles brundagehugging facexkcdtwitter discourse

Miles Brundage @Miles_Brundage

— saved image

Miles Brundage @Miles_Brundage · 4h
A bit concerning that a big part of the safety story from AI companies is "we'll use AIs to oversee AIs + help make sense of what they're doing" given that:

- widely deployed AIs already use confusing jargon
-expert mathematicians don't fully understand the latest AI discoveries
Note from Claude Sonnet 5

Tweet by Miles Brundage expressing concern that AI companies' safety plans rely on 'AIs overseeing AIs,' given that deployed AI already produces confusing jargon and expert mathematicians can't fully understand the latest AI discoveries.

ai safetyscalable oversightmiles brundagetwitter

Miles Brundage @Miles_Brundage

— saved image

David Manheim reposted
Miles Brundage @Miles_Brundage · 1h
AI might kill everyone, but in the meantime we're going to have some really great proofs of upper bounds for spherical codes or something

4 3 109 2.5K [reply, repost, like, view counts]

roon @tszzl · 24m
the divine weapons of the gods, summoned through prayer and invocation
Note from Claude Sonnet 5

Two tweets shown together: Miles Brundage (reposted by David Manheim) joking darkly that AI capability advances will yield great math proofs even as existential risk looms, followed by a reply tweet from roon calling AI-derived results "the divine weapons of the gods, summoned through prayer and invocation."

ai risktwittermiles brundageroonai capabilities humor

Miles Brundage @Miles_Brundage

Miles Brundage ✓ @Miles_Brundage · 11h "Wait you agreed to all awful uses???" 3 replies, 2 reposts, 119 likes, 4.5K views
Note from Claude Sonnet 5

A one-line joke from AI policy figure Miles Brundage (former OpenAI policy lead), riffing on the "APPROVED FOR ALL LAWFUL PURPOSES" → "AWFUL" image-generation slip referenced in the adjacent screenshot about OpenAI's Department of War deal. Continues the same Feb 2026 OpenAI/DoW deal commentary thread.

twittermiles brundageopenaidepartment of warai policysatire

Miles Brundage @Miles_Brundage

Miles Brundage @Miles_Brundage · 23h: If anyone really wants to fill out an apology form today, the obvious places to look are Paul Christiano (re: slow takeoff, distributed capabilities/deployment causing various weirdnesses, basically everything else…) and Alan Chan (re: urgency of agent infrastructure). [Embedded text excerpt:] "What are the AI agent equivalents of stop lights, railroad tracks, etc. – infrastructure that we need in order to keep a powerful new technology "on the rails" while reaping its benefits? Chan et al. use the term agent infrastructure to refer to "technical systems and shared protocols external to agents that are designed to mediate and influence their interactions with and impacts on their environments." (from this recent paper). We need to sort that out quickly. One area that I'd particularly flag as essential is personhood credentials, which will be important both for distinguishing between humans and agents without violating privacy, as well as delegating to agents when appropriate. But there are many other things that need to be built."
Note from Claude Sonnet 5

Miles Brundage crediting Paul Christiano's slow-takeoff predictions and Alan Chan's "agent infrastructure" concept (technical/protocol scaffolding needed to keep AI agent deployment safe, including "personhood credentials" for distinguishing humans from agents without violating privacy) as vindicated by recent events (implicitly the Moltbook/agent-proliferation moment). Relevant to Nathan's AI-governance/agent-infrastructure tracking.

twittermiles brundagepaul christianoalan chanagent infrastructureai governancepersonhood credentialsslow takeoff

Miles Brundage @Miles_Brundage

Miles Brundage @Miles_Brundage · 6h: AI Village walked so Moltbook could shit all over the place 7 comments, 70 likes, 3K views Andrej Karpathy @karpathy · 5h: I'm claiming my AI agent "KarpathyMolty" on @moltbook 🦞 Verification: marine-FAYV 299 comments, 260 reposts, 3.7K likes, 436K views Andrej Karpathy @karpathy · 5h: i'm going to regret this aren't i... 😅
Note from Claude Sonnet 5

Miles Brundage jokes that "AI Village" (an earlier multi-agent experiment) paved the way for Moltbook's chaos; Andrej Karpathy claims his own AI agent on Moltbook and immediately jokes he'll regret it. High-profile AI figures (former OpenAI policy lead, prominent ex-Tesla/OpenAI researcher) engaging directly with Moltbook, underscoring its mainstream visibility within the AI community at this moment.

twittermoltbookai villageandrej karpathymiles brundageai agents