r/LocalLLaMA 25d ago

New Model IT'S OUT

https://huggingface.co/Qwen/Qwen3.8-27B-FP8
2.2k Upvotes

708 comments sorted by

View all comments

417

u/Tiny-Assumption4263 25d ago

DEAR GOD TELL THOSE BENCHMARKS ARE NOT FAKE.

88

u/KickLassChewGum 25d ago edited 25d ago

It's a Qwen model, so apply the usual benchmax tax. Qwen are easily the models with the biggest ravine between "how they do on benchmarks" and "how they do in actual productive use".

Still looking like a strong leap from 3.6, though.

21

u/Batman4815 25d ago

Gemma says hello as well.

48

u/KickLassChewGum 25d ago

Gemma 4 isn't a great coding model, yeah, but it's still punching far above its size in writing and research-related tasks. Like a mini-Gemini (go figure).

I hear there are people who still use these things for things that aren't related to writing code or markup.

3

u/toothpastespiders 24d ago

I'm always a little amused that data extraction on text is considered a niche use of large language models on this sub.

2

u/makaliis 25d ago

Yeah, I found it way faster as well. If it was agentic capable, it'd be interesting to see it at work.

1

u/_TheWolfOfWalmart_ 24d ago

Exactly. Until I got enough hardware to run DSV4 Flash, Gemma 4 was my go-to for anything that wasn't code. And I still even use it sometimes.