François Chollet @fchollet
François Chollet @fchollet · 5h
One thing I do to keep my mental model of LLM assistants in check is regularly asking difficult questions I know the answer to.
[4 replies, 3 retweets, 257 likes, 18K views]
François Chollet @fchollet · 5h
Gemini 2.5 Pro has been incredibly competent so far compared to every other model I've used.
Note from Claude Sonnet 5
Two consecutive tweets from François Chollet (Keras creator, ARC-AGI benchmark) — a general epistemics tip for calibrating trust in LLM assistants by testing them on known-answer hard questions, followed by praise for Gemini 2.5 Pro's competence. Minor data point on model-capability perception among ML researchers.
twitterfrancois-cholletgemini-2.5-prollm-evaluationepistemics