← Timeline

Esben Kran

@EsbenKC on X

1 capture, most recent first. Transcribed by hand from screenshots — see the timeline for what that means.

Esben Kran @EsbenKC

— saved image

Esben Kran @EsbenKC · 2h
2022: "We can create verifiable neural nets!"

2023: "Nope... But at least we can create fully interpretable AI."

2024: "Nope... But at least we can make them as benevolent, human aligned!"

2025: "Nope... But at least we can control them!"

2026: "Nope... But at least we can slow them down by training them to think honey tokens will trigger security review!"

2027: hmmm
[cut off]
Note from Claude Sonnet 5

Tweet from @EsbenKC satirizing the year-over-year retreat of AI safety ambitions, from verifiable neural nets in 2022 down to honeypot-token deception tricks in 2026, ending with an ominous 'hmmm' for 2027.

ai safetyinterpretabilityhumoralignment