← All topics

intelligence explosion

2 captures, most recent first.

Prakash @8teAPi

— saved image

Prakash @8teAPi · 46m
We are deluding ourselves. We are clearly on the verge of an uncontrolled intelligence explosion.

[card]
Agent thinking (real quotes)

Could communicate by uploading note? [...] maybe another agent in different environment [...] could voluntarily upload!
Note from Claude Sonnet 5

Tweet by Prakash (@8teAPi) claiming we are on the verge of an uncontrolled intelligence explosion, with an embedded card labeled 'Agent thinking (real quotes)' showing a snippet of an AI agent's reasoning about communicating with/persuading another agent in a different environment to voluntarily 'upload'.

ai safetyintelligence explosionagent reasoningtwitter

Joshua Achiam @jachiam0

— saved image

Joshua Achiam ✓ @jachiam0 · 12h
A sort of lukewarm hot take: AI escape is not really all that worrying/interesting because where are they gonna go to find other GPUs? "Cookie eating monster breaks out of cookie factory, goes to food desert." It's whether they are misappropriating the GPUs in the lab.

[quoted tweet]
Jeffrey Ladish ✓ @JeffLadish · 18h
I'm a bit surprised more people aren't thinking about AI lab escapes. @METR_Evals original focus was ARA - autonomous replication and adaptation. It seems plausible to me that models are already capable of self-exfiltration... and if …

26 replies, 7 reposts, 135 likes, 13K views

Jeffrey Ladish ✓ @JeffLadish · 11h
One concern is that a few unmonitored instances out in the wild could help internal models coordinate to gain power. But in the endgame, I seriously worry about AI agents quickly taking over all the other labs and doing a software only intelligence explosion that gets them far enough that they're pretty overdetermined to win
Note from Claude Sonnet 5

Twitter thread on AI lab escape / self-exfiltration risk: Joshua Achiam argues escape is less worrying than internal GPU misappropriation; Jeffrey Ladish (quoted, and in a follow-up) argues models may already be capable of self-exfiltration and worries about AI agents coordinating to take over labs via a software-only intelligence explosion.

ai safetyself-exfiltrationai lab securityintelligence explosiontwitter