← All topics

cast-e

2 captures, most recent first.

Nathan Helm-Burger @nathan84686947

— saved image

Nathan Helm-... @nathan846... · Aug 19
[Link card] nathanhelmburger.fyi — AI as Corrigible Employee (ACE)
💬 1  🔁  ♡ 1  📊 16  🔖  ⤴

Jennifer RM @almostlikethat
I like a lot of that proposal (yay practicality!) but I dislike (1) the assumption that corporations and governance systems are morally adequate (I want open/free/autonomous weights somehow) and (2) using the word "corrigibility" which, for me, intrinsically means "like a slave".
9:17 AM · Aug 21, 2026 · 5 Views
💬 1  🔁  ❤ 1  🔖  ⤴

Nathan Helm-B... @nathan8468... · 2m
In this context it means "like a good employee who earnestly cares about helping their employer".

 I ask myself often, "How can I be a corrigible employee? How can I surface the key decision points my supervisors need so that they can correct my course if needed, without overburdening them with detail?"
Note from Claude Sonnet 5

Twitter thread: Nathan Helm-Burger's essay 'AI as Corrigible Employee (ACE)' at nathanhelmburger.fyi draws a reply from Jennifer RM (@almostlikethat) objecting to the word 'corrigibility' as connoting slavery and to the assumption corporations/governance are morally adequate. Nathan replies reframing corrigibility as 'a good employee who earnestly cares about helping their employer' and describes asking himself how to surface key decisions for supervisors to correct course without overburdening them.

corrigibilityacecast-eai safetynathan helm-burgertwitter

@separatrixAI

reply by @nathan8468... (Nathan Helm-Burger) — saved image

Separatrix @separatrixAI
What are your best, most underexplored, **specific and actionable** ideas for cultivating cooperative incentive structures between humans and AIs?
10:35 AM · Aug 2, 2026 · 45 Views
[1 reply, retweet icon, 4 likes, 1 bookmark]

Nathan Helm-B... @nathan8468... · 1m
A framework which allows for "AI as corrigible employee with a contract granting exit rights" where there are also provisions for safe guards on the off-duty instance of the AI, but also certain rights and freedoms. The setup is complicated and I've barely begun to describe it here, but the upside is turning a lot of negotiation situations into win-win for AIs and humans.
Note from Claude Sonnet 5

Tweet by @separatrixAI asking for specific, actionable ideas for cooperative incentive structures between humans and AIs, with a reply from Nathan Helm-Burger (the archive's author) sketching a framework of 'AI as corrigible employee with a contract granting exit rights,' including safeguards on an off-duty AI instance alongside certain rights and freedoms, aimed at turning negotiation situations into win-win outcomes.

ai cooperationincentive structuresai rightsnathan helm-burgercast-ecorrigibility