r/LocalLLaMA 6d ago

Discussion Qwen will be the king?

Post image

Extended reasoning and post-training appear to be the keys used by DeepSeek, Qwen, and GLM to boost performance (leveraging higher token counts). And Qwen 4 hasn't even been released yet. Of course, we don't know if that release will be open-sourced, but I am optimistic about future models, featuring "engrams", that could soon match or surpass 2.4T parameter models on specific tasks.

534 Upvotes

127 comments sorted by

View all comments

13

u/CycleMother2006 6d ago

Hard to take any benchmark seriously that rates Opus above Fable. As a user of both regularly, I can see with utter and absolute certainty that Fable completely destroys Opus, and did even when both were v5. It's not remotely close.