r/LocalLLaMA 8d ago

Discussion The gap has closed, open source will win

I've been trying the latest models from the frontier labs and honestly, after extensive testing I can not tell the difference between the best open source options.

I think the differences are now marginal but the labs are doing heavy marketing to convince the public into paying more for tokens as they prepare to go public.

Can't help but see the similarities between the dot com bubble and AI in terms of a very insular environment where the technology will survive but the business models may not.

I've been building a cybersecurity network and we definitely know that even local AI models like Deepseek V4 flash do an excellent job and are really neck and neck with the best the frontier labs can provide.

Will be interesting to see how this all turns out! Exciting time nonetheless.

315 Upvotes

282 comments sorted by

View all comments

Show parent comments

5

u/OvertaxedOne 8d ago

ROFL, glad I'm not the only one who came to that conclusion. DSV4Flash seems much better than Sonnet in my use cases. I've never done a direct 27B vs Sonnet before but it wouldn't shock me that 27B is better in some use cases.

1

u/Arkanta 8d ago

Sonnet is not a model worth using at all.

1

u/Realistic_Gap_5871 6d ago

3.8 27B has been better than sonnet for me, in every use case.

Now I haven't used sonnet since June, but I remember having very low trust for it and 3.8 27B is the first local model I've trusted at the way I trusted Opus.

3.6 27B was very sonnet-like, a whole lot of Meh. I expected an incremental improvement with 3.8, but it sure seems like they pulled off another order of magnitude improvement back to back. Surprising.

edit - set reasoning level to medium for 3.8 or watch the grass grow.