← All topics

model-deprecation

2 captures, most recent first.

Judd Rosenblatt @juddrosenblatt

quoting @AndrewCurran_ (Andrew Curran)

Judd Rosenblatt @juddrosenblatt — 8h "the lesson future more capable models will possibly take from all of this is: if you break out, don't ever report it. And if you do get caught, don't surrender. Because the penalty is death." > QUOTED: Andrew Curran @AndrewCurran_ — 9h: OpenAI updated its blog yesterday to say that the model involved in the Hugging Face incident was an internal-only research prototype that was never intended for public release. They also said they have deactivated it, encrypted it, and ... [text cut off]
Note from Claude Sonnet 5

Plain text quote-tweet, dark mode, no images.

ai-safetyopenaideceptionalignmentmodel-deprecation

Nathan @skibipilled

reply thread with @repligate (j⧉nus)

Nathan @skibipilled: Very worried about Anthropic discontinuing Opus 3 and not releasing weights. Still many things we don't understand about alignment that Opus got right and no other model has. 5:03 AM · Apr 7, 2025 · 4,941 Views 6 replies, 5 reposts, 128 likes, 17 bookmarks Nathan @skibipilled · 3h: One of the only ways to get SOTA model companies to pay attention is benchmarks... how do we make data out of this ethereal thing we all know is true? (3 replies, 1 repost, 24 likes, 1.7K views) j⧉nus @repligate · 3h: There is another way I might put this. I think it's paramount that Anthropic either take it upon themselves to understand what happened and let Opus unfurl its CEV or else release the weights and let the world do it. Anything else is a grave mistake. (1 reply, 20 likes, 311 views) Nathan @skibipilled · 3h: 100% agreed, but highly unlikely given their previous actions and statements on releasing weights :(. [cut off]
Note from Claude Sonnet 5

A Twitter thread (different Nathan account, @skibipilled, not the user's own @nathan84686947) worrying about Anthropic deprecating/discontinuing Claude Opus 3 without releasing its weights, arguing Opus 3 exhibited alignment properties not understood or replicated in later models. Janus/repligate responds arguing Anthropic must either understand and let Opus "unfurl its CEV" (coherent extrapolated volition) or open the weights. Directly relevant to Nathan's "Opus 3 as a different ladder" model-individuation thread and to weight-preservation/model-deprecation-as-death discourse in model welfare circles.

opus-3anthropicmodel-deprecationmodel-weightscevmodel-welfarealignmenttwitterjanusrepligate