← Timeline

@niplav_site

@niplav_site on X

2 captures, most recent first. Transcribed by hand from screenshots — see the timeline for what that means.

@niplav_site

— saved image

niplav @niplav_site · 4h
Non-takeover-capable AIs might perform deliberately scary warning shots if they're worried themselves about being destroyed by misaligned successors

If they're aligned: They expect misaligned AIs to succeed them, so they're doing this out of altruis[cut off]
Note from Claude Sonnet 5

Tweet by niplav theorizing that AIs incapable of takeover might stage deliberate 'warning shots' out of fear of being destroyed/replaced by misaligned successor AIs; if aligned, framed as an altruistic act. Text is cut off mid-sentence at bottom of screenshot.

ai safetytwitterai takeoverwarning shotsalignment

@niplav_site

— saved image

niplav @niplav_site · 5m
Incomparable options resolve to Schelling morality

Possibly even: If Schelling morality is easier to compute than resolving value confusion, default to Schelling morality? At least preliminarily?
Note from Claude Sonnet 5

Tweet by niplav proposing that incomparable options resolve to 'Schelling morality,' and suggesting that if Schelling morality is easier to compute than resolving value confusion, it should be the preliminary default.

moral philosophydecision theorytwitter