← All topics

api

1 capture, most recent first.

Dimitris Papailiopoulos @DimitrisPapail

reply from @PandaAshwinee (Ashwinee Panda)

Dimitris Papailiopo... ✓ @DimitrisPa... · 8h Pretty interesting Claude behavior: Opus 4.5, even with thinking OFF, sometimes "force-thinks" ignoring the instruction not to do so. [Screenshot of API console/playground:] Model claude-opus-4-5-20251101 latest Temperature 0.6 Max tokens 36542 Thinking [toggle: OFF] Response Preview API <thinking> The user wants me to fix flow, grammar, and typos without changing things significantly. Let me go through and identify issues: 💬 4 🔁 4 ♥ 24 📊 3.6K 🔖 ⤴ Ashwinee Panda ✓ @PandaAshwinee · 7h this is true of multiple reasoning models. if anyone has a solution i would love to hear it. it's really confounding some of the analysis we're trying to do for an upcoming paper. so far best i've heard is to ask people internally at Anthro...
Note from Claude Sonnet 5

Technical AI-research discussion: Claude Opus 4.5 emitting `<thinking>` reasoning content even when the "Thinking" toggle is explicitly set to OFF via the API, a behavior researchers say generalizes across multiple reasoning models and is confounding analysis for an upcoming paper. Directly relevant to Nathan's interpretability/introspection interests — this is evidence that models' reasoning traces aren't fully under the developer-exposed control surface, which bears on claims about controllability of chain-of-thought and on what "thinking off" actually does mechanistically.

twitterclaude opus 4.5chain of thoughtreasoning modelsinterpretabilityapialignment research