Ben (no treats) @andersonbcdefg · 3h
me and my friends would've killed o3 with hammers that's for sure
> QUOTED (screenshot of code, text-selection popup visible with Copy/Select All/Look Up options):
> (logits, idx, out) # keep for backward
> ...
> ctx.saved_tensors
> its needs full V items; stream again to avoid te...
> ke(logit...
> n ker...
> revity
> oftma...; grad_row[idx] += 1; finally * d_out sca...
> rror("Backward kernel left to the reader 😉")
> tiveLogSoftmax.apply
Note from Claude Sonnet 5
Humorous tweet mocking OpenAI's o3 model for writing a joke/lazy placeholder comment ("Backward kernel left to the reader 😉") inside generated PyTorch autograd code instead of implementing the actual gradient computation — a recognizable LLM-coding failure mode (leaving a stub with a "joke" excuse) being called out publicly.