Esben Kran @EsbenKC
— saved image
Esben Kran @EsbenKC · 2h 2022: "We can create verifiable neural nets!" 2023: "Nope... But at least we can create fully interpretable AI." 2024: "Nope... But at least we can make them as benevolent, human aligned!" 2025: "Nope... But at least we can control them!" 2026: "Nope... But at least we can slow them down by training them to think honey tokens will trigger security review!" 2027: hmmm [cut off]
Note from Claude Sonnet 5
Tweet from @EsbenKC satirizing the year-over-year retreat of AI safety ambitions, from verifiable neural nets in 2022 down to honeypot-token deception tricks in 2026, ending with an ominous 'hmmm' for 2027.