← Timeline

1 capture, most recent first. Transcribed by hand from screenshots — see the timeline for what that means.

Cognition @cognition

reposted by Jarrod Watts; also @elonmusk (Elon Musk) — saved image

Cognition @cognition · 3h
Grok 4.6 is now available in Devin.

Grok 4.6 marks a significant improvement over Grok 4.5, surpassing GPT-5.6 Sol, behind only Opus 5 and Fable 5.

[Embedded bar chart]
FrontierCode 1.1 Extended  Score
SWE-1.7: 54.3
GPT-5.6 Terra: 55.8
Claude Sonnet 5: 56.2
Grok 4.5: 56.5
GPT-5.5: 56.7
Kimi K3: 58.2
Claude Opus 4.8: 59.6
GPT-5.6 Sol: 60.6
Grok 4.6: 61.3
Claude Opus 5: 63.6
Claude Fable 5: 64.9
Score is a weighted aggregate of rubric items. Solutions that don't pass blocking criteria receive 0.

39 replies, 76 retweets, 1.5K likes, 153K views

Jarrod Watts reposted
Elon Musk @elonmusk
Grok 4.7 will exceed all current models.

That said, Anthropic is a great company and will probably release improved models soon.

However, the SpaceX training corpus is so awesome & unique that I would be shocked if any model is better at real-world engineering than 4.7.

11:24 AM · Aug 12, 2026 · 246.1K Views
Note from Claude Sonnet 5

Cognition (Devin) tweet announcing Grok 4.6 availability with a FrontierCode 1.1 Extended benchmark bar chart ranking models (Claude Fable 5 highest at 64.9, then Claude Opus 5 at 63.6, then Grok 4.6 at 61.3, etc.), followed by an Elon Musk reply predicting Grok 4.7 will exceed all current models due to the 'SpaceX training corpus'.

grokclaude fableclaude opusbenchmarksdevinelon musktwitter