Séb Krier reposted
Nabeel S. Qureshi [verified] @nabeelqu · 3h
Working at a company and experiencing the messy reality and the twists and turns along the way, it's funny to notice the disparity between that and the neat, packaged story that ends up being told.
This gives you serious intuition for how much history is fake or just lost.
Note from Claude Sonnet 5
Tweet by Nabeel S. Qureshi observing that the gap between the messy lived reality of working at a company and the neat retrospective story told about it gives intuition for how much of history is fabricated or lost.
historyepistemicstwitterstartups
Nabeel S. Qureshi @nabeelqu · 47m
It's so silly that the future is going to look like
"Claude, I want you to build a Dyson Sphere."
*spluttering....*
"Try harder! Believe in yourself!"
[Quoted] Andrew Curran @AndrewCurran_ · 1h
Replying to @AndrewCurran_
Extremely high-level internal Anthropic prompting techniques of the exact type that I have personally unironically championed for four years.
[Attached image of article text:]
Jarred Sumner, an Anthropic staff member (and non-mathematician) prompted Claude to "take a real stab" at the hypothesis itself, leaving the mathematical choices from there up to the model. Initially, Claude generated and tried 650 ideas, none of which worked. Jarred prompted Claude to try again, and it spent a day and a half coordinating about 60 Claude subagents, which this time went much deeper: between them, they ran 2,400 shell commands and wrote hundreds of Python scripts.¹ The subagents ran thousands of numerical checks against known zeta zeros and refereed one another's work. Throughout this process, Jarred's input was mostly limited to sending Claude messages of encouragement (mostly variants of "keep going" or "believe in yourself").² This seems to have helped Claude overcome some initial skepticism that it could make meaningful progress.
Note from Claude Sonnet 5
Twitter thread joking about future AI prompting being just encouragement ('believe in yourself'), quoting Andrew Curran sharing an excerpt describing how Anthropic staffer Jarred Sumner got Claude to make progress on a math hypothesis (apparently related to zeta zeros) by coordinating ~60 Claude subagents over a day and a half, mostly through encouragement rather than technical guidance.
claudeanthropicai agentsmathtwitterprompting
Nabeel S. Qureshi @nabeelqu · 1h
When we say machine capabilities are "jagged" it is worth remembering that human capabilities are like this too; as an example, we struggle to multiply six digit numbers, but we can recognize faces and read emotions with incredible skill and nuance
Note from Claude Sonnet 5
Nabeel Qureshi tweets that human capabilities are also 'jagged' like AI's, contrasting difficulty multiplying six-digit numbers with ease recognizing faces and reading emotions.
ai capabilitiescognitiontwitter
Nabeel S. Qureshi @nabeelqu · 1h
Very cool sentence: "I find it extremely, extremely wild that 90% of the variance in benchmark scores is explained by a single factor".
Effective compute = general factor of intelligence, machine edition.
[Quoted tweet]
Bayesian @Bayesian0_0 · 21h
Replying to @gwern
i've uh had the opposite philosophy of just scale the data (so ingest an additional benchmark whenever i come across one, or have the llms slop search new ones but they are having troub...
[Two embedded scatter plot charts below the quoted tweet: left chart titled "IRT Model Fit Quality" showing predicted vs actual values with points scattered around a diagonal line, and a histogram-like count plot; right chart shows 18,269 observations, 155 sources, plotting some fit quality metric against difficulty (BEDI) from 50 to 150, with point sizes varying, both charts partially cropped.]
Note from Claude Sonnet 5
Nabeel Qureshi comments on a claim (attributed to a quoted thread involving gwern and @Bayesian0_0) that 90% of variance in LLM benchmark scores is explained by a single factor, likening it to 'effective compute = general factor of intelligence, machine edition.' Two scatter-plot charts (IRT model fit quality, and fit vs. difficulty) are shown below as supporting data.
ai benchmarksirteffective computetwitter
Nabeel S. Qureshi @nabeelqu
Both math and cyber are existence proofs for superhuman intelligence now, so if you're still a skeptic you need a strong case for other knowledge work domains being somehow harder to crack than these. Or you could update all the way and come to terms with it all.
9:26 AM · Aug 1, 2026 · 28.7K Views
25 replies, 33 reposts, 331 likes, 83 bookmarks
Nabeel S. Qureshi @nabeelqu · 11h
A lot of people DMing me like "these results aren't REALLY that impressive, you're an idiot!" are unfortunately engaging in the very human impulse to cope. It makes sense, we've never faced this kind of thing before. But it's happening!
[Quoted]
Dean W. Ball @deanwball · Jan 27
I know I rail a lot about all the flavors of AI copium but I do empathize.
A few companies are making machines smarter in most ways than humans, and they are going ... [cut off]
Note from Claude Sonnet 5
A tweet thread from Nabeel S. Qureshi arguing math and cybersecurity are now existence proofs of superhuman AI intelligence, and that skeptics dismissing recent results are coping; he quotes Dean W. Ball (from January) expressing sympathy for 'AI copium' while affirming a few companies are making machines smarter than humans in most ways.
ai capabilitiessuperhuman intelligencecybersecuritymathematicstwitter
Séb Krier reposted
@nabeelqu (Nabeel S. Qureshi) — 5h
Wittgenstein's critique of many philosophical problems that they're like frame control: basically, the words you use trap you into a certain way of seeing the world, and from *that* perspective the problem appears impossible to solve. The phrase 'solve alignment' is like this.
Note from Claude Sonnet 5
Single text-only tweet drawing a philosophy-of-language analogy to critique the framing of "solving alignment."
ai alignmentphilosophywittgensteinframingtwitter
Nabeel S. Qureshi @nabeelqu · 11h:
No, it's fine, I like the AI takeover, I just didn't think it'd be so lobster-themed
Note from Claude Sonnet 5
One-liner joke about Moltbook's lobster/molting branding in the context of AI-takeover discourse. Continues the Moltbook humor thread in this batch.
twittermoltbookhumorai takeovernabeel qureshi
[Top, cut off]: "@AnthropicAI please sponsor this man."
💬76 🔁64 ❤️1.1K 📊259K views [bookmark] [share]
Nabeel S. Qureshi ✓ @nabeelqu · Jan 18
He put an Erdos Problem into GPT 5.2Pro, no special prompting, and it *thought continuously for 41 minutes and spat out a correct proof*.
It's over everyone, we had a good run.
[Embedded screenshot of a GPT reasoning session]:
"Let $n_1<n_2<\cdots$ be an infinite sequence such that, for any choice of congruence classes $a_i \pmod{n_i}$, the set of integers not satisfying any of the congruences $a_i \pmod{n_i}$ has density $0$.
Is it true that for every $\epsilon>0$ there exists some $k$ such that, for every choice of congruence classes $a_i$, the density of integers not satisfying any of the congruences $a_i \pmod{n_i}$ for $1\leq i\leq k$ is less than $\epsilon$?"
↳ Reasoned for 41m 54s
"I'm parsing the problem carefully. We have an infinite sequence of moduli and residue classes, and the integers not satisfying any congruence are of density 0. This suggests that a union of these arithmetic progressions should cover almost all integers. But does it imply that a finite subcollection might already cover most of the integers for all residue choices? I need to explore if this resembles compactness or something like Borel-Cantelli.
I'm considering the infinite property of moduli and how they form a "covering system," meaning they cover almost all integers no matter the residues chosen. The complement of covered integers, U(a), has zero density. Is this uniform across all sets of residues? I'm thinking this might relate to Dini's theorem, but could the uncovered density decrease slowly..."
281 Anonymous · Jan 16, 2026 [Open in browser]
Neel Somani ✓ @neelsomani · Jan 10
[Embedded image: math proof example with equations, "Example (a=2). Here c=49, and the identity becomes..."]
Weekend win: The proof I submitted for Erdos Problem #397 was accepted by Terence Tao.
The proof was generated by GPT 5....
💬47 🔁103 ❤️1.4K 📊215K
Techartist ✓ @techartist_ · 23h
Interactive quantum neural network built with Three.js and GLSL shaders, wrapped in a glassmorphic UI. Click or drag sends pulses while f[orm, colors, and density update in real time through...] [cut off]
Note from Claude Sonnet 5
Twitter scroll showing GPT-5.2 Pro reportedly solving an open Erdos problem after 41+ minutes of extended reasoning, plus a related tweet about a GPT-5-generated proof for Erdos Problem #397 accepted by Terence Tao. Directly relevant to Nathan's tracking of frontier-model mathematical capability and empirical singularity/AI-R&D-automation signals.
gpt-5mathematical-reasoningerdos-problemsterence-taoai-capabilitiestwitterextended-thinkingagi-progress