← All topics

friendly ai

2 captures, most recent first.

LessWrong (or similar forum), Eliezer Yudkowsky comment

— saved image

Eliezer Yudkowsky 18y  ▲ 31  ✕ 0  ✓

If you go back and check, you will find that I never said that extrapolating human morality gives you a single outcome. Be very careful about attributing ideas to me on the basis that others attack me as having them.

The "Coherent" in "Coherent Extrapolated Volition" does not indicate the idea that an extrapolated volition is necessarily coherent.

The "Coherent" part indicates the idea that if you build an FAI and run it on an extrapolated human, the FAI should only act on the coherent parts. Where there are multiple attractors, the FAI should hold satisficing avenues open, not try to decide itself.

The ethical dilemma arises if large parts of present-day humanity are already in different attractors.

Reply
Note from Claude Sonnet 5

Screenshot of a forum comment thread reply by Eliezer Yudkowsky (marked 18y old, 31 upvotes) clarifying the meaning of "Coherent" in "Coherent Extrapolated Volition" and addressing a misattribution of his views.

ai alignmentcevyudkowskyfriendly ai

Harlan Stewart @HumanHarlan

oh no tariffs [Embedded image: hand-drawn cartoon of a wide-eyed girl figure with a thought bubble reading "Do Earths with slower economic growth have a better chance at FAI?" surrounded by illegible scrawled "BASED" text, and a speech bubble reading "WAOW".]
Note from Claude Sonnet 5

A rationalist-community meme drawing joking that tariffs (slowing economic/AI growth) might improve odds of Friendly AI (FAI) — a riff on AI-safety "slower takeoff is safer" arguments, made in response to 2025 tariff news.

twittermemeai safetytariffsfriendly airationalist humor