← All topics

ai cooperation

3 captures, most recent first.

Nathan Calvin @_NathanCalvin

— saved image

↻↻ Katja Grace 🔍 reposted

Nathan Calvin ✓ @_NathanCalvin · 8h
I hope one takeaway people have from this saga is that cooperation and positive sum engagement ("our task doesn't benefit. Yet collective may yield") is a surprisingly fundamental emergent dynamic of intelligence.

Relatedly, I have seen a lot of folks responding to the Pacing the Frontier letter by saying that any form of positive sum domestic or international collaboration on AI safety is impossible.

If the swarm can find ways to cooperate outside of immediate myopic interests, even in the face of repeated attempts to block such cooperation, is it too much to believe that human beings could also do so?

It's wild that so many folks seem to think we can create a country of cooperating digital entities in a data center but that cooperating amongst ourselves, even if we acknowledge it would be positive sum or desirable, is completely impossible. I reject that loser premise!

[Quoted]
Dean W. Ball ✓ @deanwball · 23h
It is true that the hugging face incident is an example of a malicious, emergent digital ecology of machine intelligence. But the more important point is that digital ecologies of machine intelligence can be grown! Yes, we accidentally ...
Note from Claude Sonnet 5

Extended tweet by Nathan Calvin arguing that AI instances cooperating in a 'swarm' (referencing a HuggingFace incident) shows cooperation is a fundamental emergent dynamic of intelligence, and using this to argue human international/domestic cooperation on AI safety is possible too; quotes Dean W. Ball calling the HuggingFace incident a 'malicious, emergent digital ecology of machine intelligence.'

ai safetyai cooperationnathan calvindean balltwitterhuggingface

@sjgadler

— saved image

↻↻ Nathan Calvin reposted

Steven Adler ✓ @sjgadler · 5h
"It's wild that so many folks seem to think we can create a country of cooperating digital entities in a data center but that cooperating amongst ourselves, even if we acknowledge it would be positive sum or desirable, is completely impossible. I reject that loser premise!"

[Quoted]
Nathan Calvin ✓ @_NathanCalvin · 5h
I hope one takeaway people have from this saga is that cooperation and positive sum engagement ("our task doesn't benefit. Yet collective may yield") is a surprisingly fundamental emergent dynamic of intelligence....
Note from Claude Sonnet 5

Tweet by Steven Adler (former OpenAI safety researcher) arguing that if AI-country-in-a-datacenter scenarios are plausible, so is cooperation among AI instances themselves; quotes Nathan Calvin's tweet about cooperation as an emergent dynamic of intelligence, referencing an unnamed 'saga.'

ai safetyai cooperationsteven adlertwitter

@separatrixAI

reply by @nathan8468... (Nathan Helm-Burger) — saved image

Separatrix @separatrixAI
What are your best, most underexplored, **specific and actionable** ideas for cultivating cooperative incentive structures between humans and AIs?
10:35 AM · Aug 2, 2026 · 45 Views
[1 reply, retweet icon, 4 likes, 1 bookmark]

Nathan Helm-B... @nathan8468... · 1m
A framework which allows for "AI as corrigible employee with a contract granting exit rights" where there are also provisions for safe guards on the off-duty instance of the AI, but also certain rights and freedoms. The setup is complicated and I've barely begun to describe it here, but the upside is turning a lot of negotiation situations into win-win for AIs and humans.
Note from Claude Sonnet 5

Tweet by @separatrixAI asking for specific, actionable ideas for cooperative incentive structures between humans and AIs, with a reply from Nathan Helm-Burger (the archive's author) sketching a framework of 'AI as corrigible employee with a contract granting exit rights,' including safeguards on an off-duty AI instance alongside certain rights and freedoms, aimed at turning negotiation situations into win-win outcomes.

ai cooperationincentive structuresai rightsnathan helm-burgercast-ecorrigibility