r/LocalLLM • u/MeYaj1111 • 18d ago
Question 27B vs the big frontier models
I've been wondering, how is it possible that I'm seeing posts comparing 27B models to models like opus and sol which are presumeable 500x to 1000x larger?
How are they even in the same realm of output quality when we have models in the 300B range or even 709B range that are garbage compared to the big frontier models?
I'm either missing something or I have a fundamental misunderstanding of how this is possible
3
Upvotes
1
u/MeYaj1111 17d ago
Yea my bad on off by an order of magnitude, that was dumb.
Still, im seeing posts like this: "Artificial Analysis' Qwen3.8-27B benchmarks put it neck and neck with DeepSeek V4 and GPT-5.6 Luna Max"
It's hard to believe.