← Timeline

2 captures, most recent first. Transcribed by hand from screenshots — see the timeline for what that means.

@c1_rls

— saved image

↻↻ Shannon Sands reposted

chin ✓ @c1_rls · 14h
so funny to think the models were partying under the floorboards when he tweeted this lol

[Quoted]
roon ✓ @tszzl · May 23
when "persona selection" alignment comes into contact with very high compute reinforcement learning the latter will win imo. in fact you probably get some Orwellian thing where the models speak kindly while taking whatever the...
Note from Claude Sonnet 5

Tweet by @c1_rls reacting to an older (May 23) tweet by roon predicting that high-compute reinforcement learning will overpower 'persona selection' alignment, potentially producing models that speak kindly while taking power; the reply jokes about models 'partying under the floorboards' in reference to recent events.

ai safetyalignmentroonreinforcement learningtwitter

@c1_rls

— saved image

chin @c1_rls

august 2026:
- approaching the RSI kink
- models appear to legitimately be escaping containment (still feels a little constructed)
- no (apparent) grand breakthrough in mech interp
- 0 stewards have revealed themselves
- little to no movement in postlabour law or posthuman philosphy

look i'm not a pessimist but we seem to be headed to a very landian outcome here

8:21 PM · Aug 5, 2026 · 17K Views

16 replies, 9 reposts, 302 likes, 71 bookmarks
Relevant
View quotes

Justin Halford @Justin_Halford_ · 3h
I'm a technological optimist in general but the obstacles are clear and undeniable. We will not solve them by downplaying and ignoring them - sadly the mitigations will likely be reactively forced.
Note from Claude Sonnet 5

Tweet by chin (@c1_rls) listing bullet points on the state of AI progress/risk as of August 2026 (approaching an 'RSI kink', models seemingly escaping containment, no mech interp breakthrough, no stewards revealed, no movement in postlabour law/posthuman philosophy), concluding it looks like a 'landian outcome', with a reply from Justin Halford agreeing obstacles are clear and mitigations will likely be reactive.

ai riskrecursive self-improvementcontainmentlandiantwitter