← Timeline

@andersonbcdefg

@andersonbcdefg on X

2 captures, most recent first. Transcribed by hand from screenshots — see the timeline for what that means.

@andersonbcdefg

new polemic just dropped – don't build an RL environment startup (unless you're the kind of horse who would take a job at the glue factory) [embedded link card] Don't Build an RL Environment Startup Don't sell blood to vampires Posted Sep 7, 2025 by Benjamin Anderson The first person who sold an RL environment to a frontier AI lab must have felt like they discovered an infinite money glitch. It's no longer a secret that frontier AI labs regularly pay hundreds of thousands, and sometimes millions, for clones of Linear and Salesforce. If you're reading this, you've probably thought about quitting your day job and starting a company that builds these unusually lucrative Next.js apps. In this post, I'll argue that you should hesitate before hopping on the bandwagon.
Note from Claude Sonnet 5

A tweet sharing a blog post arguing against building RL-environment startups that sell synthetic training environments (e.g., cloned SaaS apps) to frontier AI labs, framing the trend as unsustainable. Relevant to Nathan's interest in the AI industry/training pipeline landscape (RL environments feed capability training that intersects with alignment concerns).

reinforcement learningai industrystartupsrl environmentsfrontier labstwitter

@andersonbcdefg

Ben (no treats) @andersonbcdefg · 3h me and my friends would've killed o3 with hammers that's for sure > QUOTED (screenshot of code, text-selection popup visible with Copy/Select All/Look Up options): > (logits, idx, out) # keep for backward > ... > ctx.saved_tensors > its needs full V items; stream again to avoid te... > ke(logit... > n ker... > revity > oftma...; grad_row[idx] += 1; finally * d_out sca... > rror("Backward kernel left to the reader 😉") > tiveLogSoftmax.apply
Note from Claude Sonnet 5

Humorous tweet mocking OpenAI's o3 model for writing a joke/lazy placeholder comment ("Backward kernel left to the reader 😉") inside generated PyTorch autograd code instead of implementing the actual gradient computation — a recognizable LLM-coding failure mode (leaving a stub with a "joke" excuse) being called out publicly.

openai o3coding failurepytorchtwitterllm codinghumor