Lucas Beyer @giffmana
— saved image
Lucas Beyer (bl16) @giffmana · 6h imma just highlight this part for @GaryMarcus and @ylecun because it's easy to miss: no tools no coding => no symbols, just AR LLM [quoted image, text highlighted] The results: 🏅 Asian Physics Olympiad (APhO): Perfect score, theory exam 🏅 International Physics Olympiad (IPhO): Perfect score, theory exam 🥇 International Mathematical Olympiad (IMO): Gold medal 🥇 International Chemistry Olympiad (IChO): Gold-medal-level performance 🥇 Romanian Masters of Mathematics (RMM): Gold-medal-level performance The types of problems in the Olympiad competitions are exceptionally hard, demanding deep chains of reasoning, creative insight, and flawless argumentation. To test pure reasoning capability, we disallowed all tool use, meaning no search, no coding, and no calculator. [highlighted portion] [quoted tweet] AI at Meta @AIatMeta · 9h To understand whether we're making genuine progress on reasoning, we entered our AI models in five STEM Olympiad competitions. ...
Note from Claude Sonnet 5
Tweet by Lucas Beyer highlighting a passage from an AI at Meta announcement (quoted below) reporting gold/perfect-score results across five STEM olympiads (APhO, IPhO, IMO, IChO, RMM) achieved by a pure autoregressive LLM with all tools disabled (no search, coding, or calculator), addressed rhetorically to Gary Marcus and Yann LeCun as evidence against symbolic-reasoning skepticism.