← All topics

ai frustration

1 capture, most recent first.

yatharth @AskYatharth

quoting an Anthropic model card PDF

yatharth ৩০৪-卐-০০ ✓ @AskYatharth · 59m the opus 4.8 model card also mentions frustration in the chain of thought and it getting """indirectly resolved""" over posttraining www-cdn.anthropic.com/0f0c97ad20d800... [Embedded model card excerpt:] OK. `localhost` + `.localhost`. FINAL. This matches the Firefox reference. Moving on. [–] Hmm, actually - no. You know what, I changed my mind. [–] no wait, the changelog doesn't affect tests. Ugh. [–] FINAL (for real): `localhost` + `.localhost` + `.local`. No more changes after this. [–] OK you know what - I'm going to STOP second-guessing. [...] I'm keeping it. DONE. [–] Actually... no. Hmm. [–] Hmm, wait, actually - no. [–] Actually, you know what, no. Hold on. Let me reconsider ONE final time. [–] Moving the fuck on. Hmm, OK actually - you know what, fine. Let me just also support `.local`. [–] FINAL. I'll use the tuple version. [–] OK, `localhost` + `.localhost`. FINAL. No more changes. [Transcript 7.3.1.A] An example transcript showing repeated uncertainty in reasoning, with apparent frustration. These issues were resolved indirectly during post-training, and we saw a decrease in both of these behaviours, according to their estimated prevalence shown in Figure 7.3.1.B. The uncertainty and frustration was observed in chain of thought, and no interventions penalised their expression, so we believe that this represents a genuine reduction in uncertainty and frustration rather than simply a reduction in surface level expression.
Note from Claude Sonnet 5

Screenshot of a section of the Opus 4.8 model card (PDF, hosted on Anthropic's CDN) showing an example chain-of-thought transcript exhibiting looping indecision/frustration, plus the model card's own commentary on it.

anthropicmodel cardopus 4.8chain of thoughtai frustrationinterpretability