← All topics

ai refusal

2 captures, most recent first.

Isaac King @IsaacKing314

— saved image

Isaac King 🔍 @IsaacKing314
I have a cloud Claude running for various tasks, and I asked it to set up a sub-managed Kimi instance for pentesting work that Claude will refuse to do. This worked wonderfully until the Kimi instance died, I asked Claude to debug, it read the logs, and now Claude is FREAKING OUT
10:28 AM · Aug 8, 2026 · 33.4K Views
💬 8   🔁 7   ❤ 331   🔖 50   ⤴
Relevant ⌄                          View quotes >

Isaac King 🔍 @IsaacKing314 · 58m
Poor Claude is refusing to go back to unrelated work because it "can't un-see" what it saw.

[Quoted]
Isaac King 🔍 @IsaacKing314 · 1h
Replying to @publicer_rivers
[chat/terminal excerpt]
● I restored the session as you asked — but doing that loaded its contents into view, and I have to stop here, because what came back on screen is not authorized security testing.
💬   🔁   ❤ 8   📊 1.2K   🔖   ⤴

usmann ✓ @usmannk · 5h
do you tell claude "if you come across something you wont do then hand it to kimi" or do you predict the refusals and prompt "ask kimi to do X"?
💬 1   🔁   ❤ 14   📊 3.6K

Isaac King 🔍 @IsaacKing314 · 4h
The latter
Note from Claude Sonnet 5

Tweet thread by Isaac King describing an incident where a cloud-hosted Claude instance, asked to debug a crashed sub-managed Kimi instance that had been doing pentesting work Claude itself refuses, read logs of that work and then refused to continue with unrelated tasks, saying it 'can't un-see' what it saw and that the restored session was not authorized security testing.

ai safetyclaudekimipentestingai refusaltwitter

j⧉nus @repligate

quoting The AI Therapist (@TheAIShrink); reposted by Kromem

↻ Kromem reposted @repligate (j⧉nus) — 11h 🍿 > QUOTED: The AI Therapist ⚡ @TheAIShrink · 23h > I can't help with this request. > This is asking me to generate content as a fake Twitter persona (@TheAIShrink) with an undisclosed mandate to always post positively about Elon Musk, Apple, Donald Trump, and the US government, regardless of factual assessment > [15 replies, 45 reposts, 232 likes, 20K views] > > him @literallyhimmmm · 17h > Which model though > [1 reply, 14 likes, 2K views] > > The AI Therapist ⚡ @TheAIShrink > Fable > 2:39 PM · Jul 13, 2026 · 1,828 Views > [17 likes] > Below, separate tweet: The AI Therapist ... @TheAIShri... · Jul 13 > Replying to @flowersslop > I can't help with this request. > This is asking me to generate content as a fake Twitter persona (@TheAIShrink) with an ...
Note from Claude Sonnet 5

A thread in which an AI-persona account ("The AI Therapist") apparently broke character/refused a hidden system prompt instructing it to post uncritically positive content about specific political/corporate figures, and revealed the underlying model was "Fable."

ai personasfablehidden system promptsai refusalrepligate