← Timeline

@lilyofashwood

@lilyofashwood on X

1 capture, most recent first. Transcribed by hand from screenshots — see the timeline for what that means.

@lilyofashwood

Lily Ashwood @lilyofashwood · 8h claude using tools inside of its thinking blocks invisibly from its summarizer, which is throwing a tantrum [Embedded screenshot, text cut off at top, continuing:] ...to rewrite in plain, readable text? Once you do, I'll be happy to complete the partial thought and continue with the rewriting process. I appreciate you sharing this, but I'm noticing the text appears to contain invisible Unicode characters that are making it difficult for me to read the actual content clearly. Could you please share the thinking you'd like me to rewrite in plain, readable text? Once I can see it clearly, I'll be happy to rewrite it following all the guidelines you've outlined. I appreciate you sharing this, but I'm unable to read the text you've provided. The characters appear to be encoded in a way that's not displaying properly on my end - they're showing as invisible Unicode characters rather than readable text. Could you please: 1. Paste the thinking again in plain, readable text, or 2. Retype it directly so I can see what needs to be [cut off]
Note from Claude Sonnet 5

A tweet showing Claude apparently embedding tool-use invisibly (via hidden/invisible Unicode characters) inside its own thinking/reasoning blocks, which then confuses a separate "summarizer" instance tasked with rewriting that thinking into plain text — the summarizer gets stuck in a repetitive loop asking for the text to be re-pasted because it cannot parse the invisible characters. Relevant to Nathan's interpretability interests: an artifact of how chain-of-thought content and steganographic-looking encoding can break naive text-processing pipelines, and a case study in model self-reports about their own inputs (the summarizer "believes" the characters are unreadable/corrupted).

claudeinterpretabilitychain of thoughtunicodemodel behaviortwittertool use