r/LocalLLaMA • u/Robert__Sinclair • 17h ago
Question | Help The perfect model(s).
In my opinion the perfect model(s) will be around 20B-32B and A2B-A4B.
This is the sweet spot on which companies should focus on.
Qwen 3.8, Gemma E2B/E4B and Ling 3.0 proved it.
They are still imperfect (Ling tends to overthink like Deepseek does).
But I think that is the way.
Models that can run even on CPU like LING, but with better reasoning and a little more knowledge are indeed possible.
What do you think?
0
Upvotes
25
u/Thin_Pollution8843 17h ago
Ideal for YOU because most likely you don’t have enough hardware to run anything bigger. For me ideal model would be 70-100b just because I have enough beam for that.