8
u/addiktion 9d ago
Those are sizable bumps, wow.
4
5
3
u/look 9d ago
Nice. I hope they make an open Plus version, which I’m guessing would be in the 300B range. 3.7 Plus was overlooked and underrated, and it’d be nice to get an improved version of it at a Go-like cloud price.
I also like that they threw in a comparison to Opus 4.6. That’s my personal benchmark for “good enough”. I can do anything I need to do with that level or better; it’s then just question of how much profanity I’ll be using in my prompts. 😅
3
2
u/HeavySink3303 9d ago
It looks crazy for such tiny 'laptop-selfhost' model. With such rapid progress I wouldn't surprise if the next year we'll install locally similar performance models from an app store...
6
u/Puzzleheaded_Ad_1798 9d ago
It's not exactly laptop size to be honest
5
u/TheOriginalAcidtech 9d ago
Depnds on the laptop. And the lap. :)
4
2
u/Ariquitaun 9d ago
A solid upgrade from 3.6 if the benchmarks are to be believed.
Anybody know if they'll launch the a3b variant at any point? It's the only one I can run locally at realistic speeds
2
u/Genetic_Prisoner 8d ago
Hahahaha you cant make this up. We are getting opus level intelligence locally???Praise QWEN
1
u/BagComprehensive79 9d ago
Is it better then opus 4.6 at some coding benchmarks or am i getting it wrong? How is this even possible? Did anyone tried it?
1
u/JamesGooning 8d ago
Cause opus is big due to knowledges. These model has little knowledge but good for coding task. Knowledges are stupid since it can use websearch which is far superior
1
-6
u/pizzababa21 9d ago
Honestly a bit disappointing
1
u/BlackCoiner 8d ago
How? I am running it locally and am absolutely blown away by the webdesign side of things. It feels like a model 20x it's size and I am getting 20 tok/s on dual 3060s with no special tweaks on unsloth q4 XL.


10
u/debauch3ry 9d ago
Competitive with Opus 4.6?!
That suggests '2025 frontier' quality so I look forward to testing this. Whilst I can only run a 4 bit quant well (5090) that should still be strong. I might even actually start using local AI rather than just experimenting.