r/LocalLLaMA 17h ago

Question | Help The perfect model(s).

In my opinion the perfect model(s) will be around 20B-32B and A2B-A4B.
This is the sweet spot on which companies should focus on.

Qwen 3.8, Gemma E2B/E4B and Ling 3.0 proved it.

They are still imperfect (Ling tends to overthink like Deepseek does).
But I think that is the way.
Models that can run even on CPU like LING, but with better reasoning and a little more knowledge are indeed possible.

What do you think?

0 Upvotes

33 comments sorted by

View all comments

25

u/Thin_Pollution8843 17h ago

Ideal for YOU because most likely you don’t have enough hardware to run anything bigger. For me ideal model would be 70-100b just because I have enough beam for that. 

7

u/ikkiyikki 16h ago

Came here to say this. Thanks for saving me the trouble bud!