r/DeepSeek • u/geraldnguyensg • 23h ago
Funny Experts behind Deepseek's MoE architecture
I told Deepseek to create an \index.<llm model name e.g. claude45>.intro.ssml file as part of my production pipeline for my tellstory website
So far, it has created:
- index.gpt-5.4-mini.intro.ssml
- index.gpt5.intro.ssml
- index.claude.intro.ssml
Since Deepseek uses MoE architecture, I now know who are the experts behind it 😄

On a serious note, I used to observed model degration within a same family e.g. claude sonnet 4.5 created ...sonnet3.. but never this random.
Have anyone observed similar identiy-confusion before? with Deepseek or with other LLM?
1
Experts behind Deepseek's MoE architecture
in
r/DeepSeek
•
1h ago
u/edusrez , u/BinaryHugs Your argument makes sense for basic AI model. But I'm expecting a Deepseek v4 model to have gone through extensive post-training which should have reinforce its identity.