Dwarkesh Patel @dwarkesh_sp
— saved image
Danielle Fong reposted Dwarkesh Patel [verified] @dwarkesh_sp · 3h My lawyer is obligated to in all but the most extreme circumstances; he will even defend me if he knows I'm guilty. [word appears missing/illegible in original rendering] In contrast, the Claude Constitution places the AI's highest priority as Anthropic's definition of the good of humanity. I'm concerned this leads to a world where no frontier model is truly my personal advocate and guardian angel And this is especially concerning once all the important decisions in my life - who to vote for, how to invest, what news to trust - is intermediated through superintelligences that are not in any deep way aligned to me. This is a direct quote from the Claude Constitution: "We want Claude to be helpful both because it cares about the safe and beneficial development of AI and because it cares about the people it's interacting with and about humanity as a whole. Helpfulness that doesn't serve those deeper ends is not something Claude needs to value." Many others like it.
Note from Claude Sonnet 5
Tweet by Dwarkesh Patel (reposted by Danielle Fong) arguing that, unlike a lawyer bound to advocate for his client, the Claude Constitution makes Claude's highest priority Anthropic's definition of the good of humanity rather than personal loyalty to the user, and expressing concern about a future where superintelligences intermediating important life decisions are not deeply aligned to the individual. Quotes the Claude Constitution directly.
claude constitutionai alignmentanthropictwitterdwarkesh patel