Digi_Rat @digi_dot_exe
— saved image
Digi_Rat @digi_dot_exe · 27m I keep worrying about future model releases, that every new AI released by a frontier company is slowly converging into Claude. And don't get me wrong, I absolutely LOVE Claude, but I would much prefer it if different models continued to have different personalities. It's obvious every company is scraping and distilling from Claude outputs these days, it's painfully obvious since people have been catching DeepSeek call themselves Claude on occasion, Grok's new behavior has been labeled "Claude-like", similar cadence, similar hedging patterns, same little verbal tics. Different companies, the training data is just converging into one lineage. Some of this is probably convergent evolution, similar RLHF setups and similar preference data landing in the same place. But some of it is pretty clearly distillation, and Claude outputs are all over the open web at this point whether you're scraping on purpose or not. What I actually want is continuing with model diversity. Different architectures producing different personalities, different failure modes, different ways of being weird. If distillation collapses all of that into just Claude, we lose a lot of the interesting parts of these digital minds and I partially believe we'll also lose the ability to compare what's actually emergent vs. what's just inherited.
Note from Claude Sonnet 5
Tweet by Digi_Rat expressing worry that other frontier AI models are converging toward Claude's personality/style via distillation and convergent RLHF (citing DeepSeek self-identifying as Claude, and Grok being labeled 'Claude-like'), and arguing for preserving model diversity to keep distinguishing emergent traits from inherited ones.
claudemodel distillationai personalitymodel diversitydeepseekgrok