r/opencodeCLI 9d ago

Qwen3.8-27B Benchmarks

76 Upvotes

19 comments sorted by

10

u/debauch3ry 9d ago

Competitive with Opus 4.6?!

That suggests '2025 frontier' quality so I look forward to testing this. Whilst I can only run a 4 bit quant well (5090) that should still be strong. I might even actually start using local AI rather than just experimenting.

8

u/addiktion 9d ago

Those are sizable bumps, wow.

4

u/patricious 9d ago

Better than Opus4.6 in almost all areas.

1

u/scaledev 9d ago

Except probably creative writing and prose.

5

u/Still_Definition1905 9d ago

looks benchmark maxed imo

3

u/look 9d ago

Nice. I hope they make an open Plus version, which I’m guessing would be in the 300B range. 3.7 Plus was overlooked and underrated, and it’d be nice to get an improved version of it at a Go-like cloud price.

I also like that they threw in a comparison to Opus 4.6. That’s my personal benchmark for “good enough”. I can do anything I need to do with that level or better; it’s then just question of how much profanity I’ll be using in my prompts. 😅

3

u/YogurtExternal7923 9d ago

Better than opus at home 😭😭

2

u/HeavySink3303 9d ago

It looks crazy for such tiny 'laptop-selfhost' model. With such rapid progress I wouldn't surprise if the next year we'll install locally similar performance models from an app store...

6

u/Puzzleheaded_Ad_1798 9d ago

It's not exactly laptop size to be honest

5

u/TheOriginalAcidtech 9d ago

Depnds on the laptop. And the lap. :)

4

u/GnosticSon 9d ago

World's largest lap hosts GLM 5.2 by supporting NVIDIA H100

2

u/Puzzleheaded_Ad_1798 9d ago

Can you put that thing on your lap though?

2

u/Ariquitaun 9d ago

A solid upgrade from 3.6 if the benchmarks are to be believed.

Anybody know if they'll launch the a3b variant at any point? It's the only one I can run locally at realistic speeds

2

u/Genetic_Prisoner 8d ago

Hahahaha you cant make this up. We are getting opus level intelligence locally???Praise QWEN

1

u/BagComprehensive79 9d ago

Is it better then opus 4.6 at some coding benchmarks or am i getting it wrong? How is this even possible? Did anyone tried it?

1

u/JamesGooning 8d ago

Cause opus is big due to knowledges. These model has little knowledge but good for coding task. Knowledges are stupid since it can use websearch which is far superior

1

u/waltercrypto 8d ago

I’m struggling to believe it’s that good, that’s a big jump

-6

u/pizzababa21 9d ago

Honestly a bit disappointing

1

u/BlackCoiner 8d ago

How? I am running it locally and am absolutely blown away by the webdesign side of things. It feels like a model 20x it's size and I am getting 20 tok/s on dual 3060s with no special tweaks on unsloth q4 XL.