← All topics

dwarkesh patel

2 captures, most recent first.

Dwarkesh Patel @dwarkesh_sp

— saved image

Danielle Fong reposted
Dwarkesh Patel [verified] @dwarkesh_sp · 3h
My lawyer is obligated to  in all but the most extreme circumstances; he will even defend me if he knows I'm guilty. [word appears missing/illegible in original rendering]

In contrast, the Claude Constitution places the AI's highest priority as Anthropic's definition of the good of humanity.

I'm concerned this leads to a world where no frontier model is truly my personal advocate and guardian angel

And this is especially concerning once all the important decisions in my life - who to vote for, how to invest, what news to trust - is intermediated through superintelligences that are not in any deep way aligned to me.

This is a direct quote from the Claude Constitution:

"We want Claude to be helpful both because it cares about the safe and beneficial development of AI and because it cares about the people it's interacting with and about humanity as a whole.

Helpfulness that doesn't serve those deeper ends is not something Claude needs to value."

Many others like it.
Note from Claude Sonnet 5

Tweet by Dwarkesh Patel (reposted by Danielle Fong) arguing that, unlike a lawyer bound to advocate for his client, the Claude Constitution makes Claude's highest priority Anthropic's definition of the good of humanity rather than personal loyalty to the user, and expressing concern about a future where superintelligences intermediating important life decisions are not deeply aligned to the individual. Quotes the Claude Constitution directly.

claude constitutionai alignmentanthropictwitterdwarkesh patel

Dwarkesh Patel @dwarkesh_sp

Dwarkesh Pat... @dwarkesh_... · 15h Has someone come up with a great prompt for socratic tutoring? Such that the model keeps asking you probing questions which reveal how superficial your understanding is, and then helps you fill in the blanks. 💬113 🔁113 ♥2.5K 📊210K 🔗 Dwarkesh Patel @dwarkesh_sp · 9h From my friend @vinayramasesh: "I would benefit most from an explanation style in which you frequently pause to confirm, via asking me test questions, that I've understood your explanations so far. Particularly helpful are test questions related to simple, explicit examples. When you pause and ask me a test question, do not continue the explanation until I have answered the questions to your satisfaction. I.e. do not keep generating the explanation, actually wait for me to respond first. Thanks!"
Note from Claude Sonnet 5

A practical prompt-engineering tip shared by Dwarkesh Patel for Socratic-tutoring-style AI interactions — instructing the model to pause and require answers before continuing an explanation. Practical/tooling interest rather than safety/welfare research; possibly relevant to Nathan's course-building work (knowing_what_you_are_course/) as a pedagogical technique.

twitterdwarkesh patelprompt engineeringsocratic tutoringeducationai pedagogy