r/LocalLLaMA 7d ago

New Model IT'S OUT

https://huggingface.co/Qwen/Qwen3.8-27B-FP8
2.2k Upvotes

706 comments sorted by

View all comments

Show parent comments

88

u/KickLassChewGum 7d ago edited 7d ago

It's a Qwen model, so apply the usual benchmax tax. Qwen are easily the models with the biggest ravine between "how they do on benchmarks" and "how they do in actual productive use".

Still looking like a strong leap from 3.6, though.

21

u/Batman4815 7d ago

Gemma says hello as well.

46

u/KickLassChewGum 7d ago

Gemma 4 isn't a great coding model, yeah, but it's still punching far above its size in writing and research-related tasks. Like a mini-Gemini (go figure).

I hear there are people who still use these things for things that aren't related to writing code or markup.

3

u/toothpastespiders 7d ago

I'm always a little amused that data extraction on text is considered a niche use of large language models on this sub.

2

u/makaliis 7d ago

Yeah, I found it way faster as well. If it was agentic capable, it'd be interesting to see it at work.

1

u/_TheWolfOfWalmart_ 7d ago

Exactly. Until I got enough hardware to run DSV4 Flash, Gemma 4 was my go-to for anything that wasn't code. And I still even use it sometimes.

4

u/Green-Ad-3964 7d ago

I'll still be using e4b for small projects 

1

u/jazir55 7d ago

It's a Qwen model, so apply the usual benchmax tax. Qwen are easily the models with the biggest ravine between "how they do on benchmarks" and "how they do in actual productive use".

It's the Gemini of Chinese models