Andrew Curran @AndrewCurran_ · 16m
This is my favorite model of all time. It felt unearthly, like talking to an alien. A spectral visitor. It made me realize this wasn't just going to be a technical revolution, but something much stranger. Everything that has happened in the last four years has been in its shadow.
Greg Brockman @gdb · 1h
GPT-4 finished training four years ago today.
Note from Claude Sonnet 5
Tweet from Andrew Curran reminiscing about GPT-4 (quote-tweeting Greg Brockman's note that GPT-4 finished training four years prior), describing it as unearthly and formative for his sense of what AI progress would mean.
gpt-4ai historyopenai
Aidan McLaugh... @aidan_mcl... · 4h
the jump from gpt4 -> gpt5 was obviously larger than the jump from gpt3 -> gpt4
[Chart, Epoch AI: "Accuracy" (y-axis 0-100%) vs "Release date" (x-axis GPT-3, '21, '22, GPT-4, '24, '25, GPT-5). Five benchmark lines: MMLU (blue, +43% GPT-3→GPT-4), TruthfulQA (teal, +40%), HumanEval (yellow, +67%), MATH (brown, +37%), GPQA Diamond (purple, +54% GPT-4→GPT-5), MATH Level 5 (orange, +75%), Mock AIME 24-25 (pink, +80%). Footnote: "*MATH Level 5 is the most difficult subset of the original MATH benchmark. Figure only includes OpenAI models."]
Note from Claude Sonnet 5
A tweet with an Epoch AI chart arguing (contra popular narrative) that GPT-4→GPT-5 benchmark gains were larger than GPT-3→GPT-4 gains, especially on hard math/reasoning benchmarks (MATH Level 5, Mock AIME). Relevant to Nathan's tracking of empirical AI capability progress/scaling trajectory (cf. his singularity-rate tracking notes, Davidson/Houlden, METR).
twittergpt-5gpt-4benchmarksepoch aicapability scalingai progress