← Timeline

Rishabh Agarwal

@agarwl_ on X

1 capture, most recent first. Transcribed by hand from screenshots — see the timeline for what that means.

Rishabh Agarwal @agarwl_

quoting @dwarkesh_sp (Dwarkesh Patel)

Rishabh Agarwal (@agarwl_) — 2h Problems we care about are often very slow verification loops (e.g, automating pretraining, making a new material)-- either you find a good enough proxy (e.g, simulator, grindable env) or unlock how to deal with this slowness (e.g, very sample efficient RL), which would be a step change. > QUOTED: Dwarkesh Patel (@dwarkesh_sp) — Jun 26 > Here's a question I find confusing and interesting and which actually tells us a lot about the nature of current AI progress: > Why has progress on computer use been so ... [truncated by platform]
Note from Claude Sonnet 5

Text-only quote-tweet chain discussing AI research/RL methodology, no images. Dwarkesh's tweet is cut off by platform truncation ("...").

ai researchreinforcement learningsample efficiencytwitter