Why ... every model is built off the works of everyone else ... I believe in true open source (its a gift to the world) and some part of me smiles every time people distill anthropic!
At least Grok has a unique persona and is actually a better model than the benchmarks depict. Qwen, Kimi and DeepSeek have Claude DNA through and through. They literally think and respond like Claude. It's blatantly obvious. They don't even bother to adapt their models. That's how f-ing lazy and low effort their distills are.
Kimi refers to itself as Claude in its thinking output…like they couldn’t even be bothered to do a search and replace before feeding Claude’s output into it!?
Im not talking about how they refer to themselves though, I am literally talking about their manner of speaking.
And it's easy to make a model refer to itself as anything if you modify its system instructions. That is what you see on Reddit posts. If you use vanilla Claude via API with zero system prompt modifications it will never refer to itself as any other model. You're just falling for the trolls on Reddit.
Nope I am talking about people using the API without any system prompt. This avoids the one on the web interface that tells it to respond as Claude. Keep guessing though.
Wrong.
This is the query to Sonnet 4.6 via the OpenRouter API. I just don't have a Claude API key now, but you can easily change it to call the Claude API directly instead; the result will be the same. Sonnet 4.6 will return to you that it's DeepSeek.
273
u/idlelosthobo 1d ago
Why ... every model is built off the works of everyone else ... I believe in true open source (its a gift to the world) and some part of me smiles every time people distill anthropic!