← All topics

career-move

2 captures, most recent first.

@ArthurConmy

Arthur Conmy ✔ @ArthurConmy I'm joining Anthropic! I'll start work on aligning upcoming models as they're trained Claude's capabilities are extraordinary. But like all models thus far, Claude isn't aligned enough to safely delegate AGI development to I can't think of a better place to work on this at 9:29 AM · Jun 24, 2026 · 148 Views 💬 1 🔁 ❤ 16 🔖 ⤴ Arthur Conmy ✔ @ArthurConmy · 4m Q: What do I mean by aligning upcoming models? A: Triaging signs of misalignment in training, then aiming for root-cause fixes over whack-a-mole patches (e.g. training against the behavior). This post is the best work I've seen on how to align models: alignment.anthropic.com/2026/teaching-... 💬 1 🔁 ❤ 4 📊 65 🔖 ⤴ Arthur Conmy ✔ @ArthurConmy · 3m I'll move from London to SF for this work! 😢
Note from Claude Sonnet 5

Three-tweet thread announcing a job move to Anthropic's alignment team, with engagement counts visible on each tweet; no images.

anthropicalignmentcareer-moveinterpretabilitytwitter

Rishub Jain @shubadubadub

reply from @JacquesThibs

Rishub Jain @shubadubadub · 3h 🧑‍🦱 After 7 years, I've just left Google Deepmind to start an AI Safety nonprofit, around Scalable and Human Oversight (i.e. building stronger "judges")! 🇨🇦 And, I'll be at FAccT in Montreal this week! (1/4 🧵) [Embedded photo: selfie in front of a "Google DeepMind" sign on an office wall.] 💬 28 🔁 19 ❤ 461 📊 28K 🔖 ⤴ Jacques ✔ @JacquesThibs · 1h FYI, we opened up an AI safety coworking space and there's an event on the FaaCT board on Saturday (unfortunately I will personally be out of town this weekend, though will be there tomorrow and maybe Friday). [Link card: horizonomega.org — "Ω Labs – HΩ"]
Note from Claude Sonnet 5

Announcement tweet with a selfie in front of a Google DeepMind office sign, plus a reply promoting an AI safety coworking space with a linked website card.

ai-safetydeepmindcareer-movefacct-conferencetwitter