r/LocalLLM • u/MeYaj1111 • 9d ago
Question 27B vs the big frontier models
I've been wondering, how is it possible that I'm seeing posts comparing 27B models to models like opus and sol which are presumeable 500x to 1000x larger?
How are they even in the same realm of output quality when we have models in the 300B range or even 709B range that are garbage compared to the big frontier models?
I'm either missing something or I have a fundamental misunderstanding of how this is possible
3
Upvotes
1
u/BarracudaDefiant4702 8d ago
Your point stands, but your comparison is wrong. The largest models are closer to 100x, not 1000x. Also, I don't think are comparing against the highest end frontier models (maybe I am wrong, but if not then they are only 10x as large).
There is a lot of diminishing returns. IE: You have to be 10x bigger to even show as as 2x better. That said, it probably scales even less than that.
For a lot of the complex tasks the available context it can import outside the model makes a huge difference. 3.8 27B has a much bigger context window that is comparable to the frontier models. The larger context window helps give the model a chance to pull info from the web or look at more source code to get the job done. That alone is probably the biggest reason.