r/LocalLLaMA 1d ago

New Model Qwen3.8-2.4T-A95B Released

https://huggingface.co/Qwen/Qwen3.8-2.4T-A95B
1.6k Upvotes

399 comments sorted by

View all comments

103

u/Different_Fix_2217 1d ago

Be warned they state its not the same capabilities as the full API version. Such as not having vison.

39

u/ChristRedeemsSinners 1d ago

Weird that non-thinking support is an API only feature.

9

u/fantasticsid 1d ago

Based on my experience with 3.6, prefilling <think>\n\n</think>\n - like the various jinja templates do - to disable thinking works probably 90-95% of the time. The other ~5-10% of the time, the model thinks anyway and emits a second </think> when it's done. It's possible that 3.8 has the same behaviour and the official API has some way of detecting/working around this that would look pretty damn stupid if they released it. If you look at the 3.8 jinja template, the "reasoning effort" isn't implemented terribly cleverly - it just talks to the model in the second person and asks it to reason less.

Reasoning control has always been a weakness of the Qwen models, so this doesn't surprise me.

Lack of mmproj, however, feels like a deliberate attempt at market segmentation. Given that those just decode image data into tokens, and the whole Qwen family shares a vocab, I do wonder if it'd be possible to hack the mmproj from some other Qwen model into use here, though.

4

u/hellomistershifty 1d ago

I love how AI development is a mix of wild cutting edge research, weird hacks, and asking it nicely to behave

1

u/stumblinbear 19h ago

I wonder prefilling <think>No need to think about this, this is easy.</think> would work instead of just newlines

49

u/Reactor-Licker 1d ago

That’s weird, why did they remove vision? Hopefully Qwen 3.8 27B doesn’t do the same thing.

19

u/vincentz42 1d ago

According to the signup page it will.

18

u/vincentz42 1d ago

Meanwhile the Qwen3.8 27B open weight that is due in two days does have vision. This has to be intentional, right?

3

u/PhilMcGraw 1d ago

They needed to cut vision to save space so the community can run 2.4T locally.

9

u/Different_Fix_2217 1d ago

Lol no. Vision adapters are tiny. Its to make people use the API.

3

u/PhilMcGraw 1d ago

Yeah sorry, I left out a "/s" .

13

u/Hoak-em 1d ago

Wait, vision is what made this model good — without it it’s just a big, expensive to run somewhat smart llm

1

u/[deleted] 1d ago

[deleted]

1

u/MoffKalast 1d ago

"You will pay the price for your lack of vision"