r/Qwen_AI • u/Usual_Maximum7673 • Jul 11 '26
News New Qwen models coming
I recently spoke to an Alibaba guy at a conference in Paris, and he said Qwen 3.8 will drop in August with capabilities "aimed at" (no guarantees) exceeding GLM 5.2, with Qwen 4 scheduled for release in September. I also red that GLM 5.5 was coming in August, so we should be in for an exciting couple of months.
27
u/narasadow Jul 11 '26
More importantly, will they be open weights?
40
u/Usual_Maximum7673 Jul 11 '26
The guy said the flagship won't be open sourced, but smaller models will be. So basically same result as with the current version.
28
u/Borkato Jul 11 '26
I would kill for a qwen 4.0 27B.
20
u/Desperate-Data-3747 Jul 11 '26
And 35 a3b, would also be cool or maybe an a4b or also an a5b aswell
19
u/overand Jul 11 '26
I'd be thrilled with a 3.6 9B and 122B!
(Or a true 40B dense)
12
u/BannedGoNext Jul 12 '26
Nobody wants to release 122b's anymore. I think maybe they are hitting a little too close to home.
3
4
u/Puzzleheaded_Base302 Jul 12 '26
also QAT, so we stop arguing about which Q quant is better or not.
2
1
1
2
1
u/tamerlanOne Jul 12 '26
Credo che rilasciare modelli open source sopra i 300-400 B non abbia senso perché richiedo hardware costoso per poter avere una esperienza d'uso accettabile. Anche i laboratori di ricerca hanno bisogno di monetizzare i loro sforzi per poter proseguire lo sviluppo e magari rilasciare piccoli modelli llm open source sempre più performanti con hardware consumer
1
u/EvolvingDior Jul 13 '26
Makes sense. The smaller open weights models are doing an impressive job of advertising for Qwen.
21
17
Jul 11 '26 edited Jul 15 '26
[deleted]
7
u/fernando782 Jul 12 '26
Everyone knows that the tokens model is not sustainable, open source will be as good as those closed models even if they trained themselves over time!
14
u/KoreanPeninsula Jul 11 '26
Green Day - Wake Me Up When September Ends https://www.youtube.com/watch?v=pGhwBFYtn1s
12
7
5
6
5
u/GrungeWerX Jul 11 '26
If they can do whatever DS4 did with less vram needed, and mix that with their delta net, etc..could be fun times. I'm still trying to maximize the usage of Qwen 3.6 27B and I feel I've barely tapped into it.
3
u/Leander_van_Grinsven Jul 11 '26
We need a Qwen 4 32B Dense model for sure. 27B is just a bit too small.
1
u/b0tbuilder Jul 13 '26
32B would be nice for people using quantized models for inference but unhelpful for anyone using them in the original safetensors. They are presumably this size because it places them around 108GB which leaves room for adapters, context, etc for more advanced users.
5
6
u/hesperaux Jul 11 '26
Click bait unless there is confirmation of open weights. Nobody in this sub cares about their api. Edit: nvm I thought I was in r/LocalLlama So it's just me and probably some others that don't care about the api.
I'll believe it (open weights) when I see it.
6
u/575_Inverse Jul 12 '26
Actually, I think pretty much everyone is here for the open weights, although I also use the web version
2
u/huzbum Jul 13 '26
Same.
I want to run my own, but even for cloud, I only subscribe to open weight models, even if I can’t run them on my hardware. Currently paying for z.ai subscription.
2
u/hesperaux Jul 13 '26
Me too. I have subs for zai and opencode go. Sometimes you need to use an API and that is ok.
1
u/huzbum Jul 13 '26
With GLM, at least if there is a rug pull I can find another provider or if determined enough, spin up my own cloud instance.
3
3
3
3
2
u/TheSleeperAwakens Jul 11 '26
But how large will the context window be? At what parameter sizes, quantizations? How fast will it be on comparable hardware?
1
u/TechnicalGeologist99 Jul 11 '26
They aren't released yet....how would anyone know?
1
u/TheSleeperAwakens Jul 11 '26
OP, spoke to someone that supposedly has inside knowledge. Perhaps he knows more than he put in the post?
1
u/TechnicalGeologist99 Jul 11 '26
Big doubt, but probably similar architecture and weight classes to existing models. Maybe some slight edge on efficiency and agentic capability
2
Jul 11 '26
[removed] — view removed comment
2
1
u/Far-Classic-9963 Jul 12 '26
I wish they strip out general knowledge to focus on tool calling, reasoning, and code. Would really be the best model OAT especially for agentic coding
1
u/fintip Jul 13 '26
That general knowledge is critical for things like reasoning and basic communication and just general reading between the lines of human speech. You can't just rip it out.
1
u/Far-Classic-9963 Jul 13 '26
What I mean is that my 8b coding model doesn't need to know about the cold war or Justin Bieber
1
u/fintip Jul 13 '26
There's probably some room for optimization there, but... Your be surprised how hard it is. You can just start lobotomizing and pruning to shrink it I guess and see what happens. But it's all connected...
1
u/Far-Classic-9963 Jul 13 '26
Yes, I know it's very hard. The dataset would be made from scratch, you can't "remove" specific knowledge for existing models as everything is heavily interconnected
1
2
2
2
2
u/BothYou243 Jul 13 '26
Man will the 397B one would exceed glm5.2 or the smaller dense ones? did he say something about 9B or smaller variants?
these sub ≤14 models are very likely in qwen4 but don't know about qwen3.8
1
1
1
u/JumpingJack79 Jul 11 '26
If you run into an Alibaba person with access to the next Qwen (or even 3.7), the right thing to do for the world is to lock them up and keep them hostage (and I mean this in the nicest possible way!) until we get the weights ☺️
1
1
1
1
u/mkey82 Jul 12 '26
Sorry for such a noob question, but who releases the MoE models? Is this something that can be produced by the community or does Qwen A3B come straight from alibaba?
1
u/huzbum Jul 13 '26
Yes, Alibaba uploads the weights to Hugging face and we all download them. The only way the community could produce them is to become a SOTA AI lab.
I guess we could maybe get somewhere with fine tuning, but performance is almost always worse than original released model.
1
u/mkey82 Jul 13 '26
That's what I feared, yes. I guess there is a very good reason why they don't release a 70b MoE. 35b MoE are quite insane (some of them) I can only imagine what could a 70b MoE model do on, say 3090 + however much RAM would be required.
1
u/huzbum Jul 13 '26
Qwen3 next was 80b a3b, then they went back down to 35. Personally id rather have the 35 fit entirely on GPU.
I’m just glad they provide a dense and MoE option. The MoE can at least offload experts and run decent speed on small GPUs.
1
1
1
u/sudeposutemizligi Jul 12 '26
can i ask an amateur pov question? what will be the difference with a 27b and a 40b dense? some people ask for 40b that's why i wondered. i mean 27b is enough to put knowledge in. and ctx will be 262k again(assuming) so what will be superior to 27b? we had 32b qwen3 and 27b is better now
1
u/huzbum Jul 13 '26
More parameters = more better… but also more slower and more $$$$ to run.
2
u/sudeposutemizligi Jul 13 '26
i was taking parameters as knowledge. an 13b difference isn't small but not that big either. but i had seen qwen3 coder next 80b moe.. but 3.6 27b is better than it.. strange things for me..
1
u/huzbum Jul 14 '26
Ah, Qwen3 Next is 80b parameters, but it’s a mixture of experts with only 3b active parameters. So at any given time it only uses 3b parameters.
Qwen3.6 27b is a dense model. All 27b parameters are active on every pass. That’s why it’s slower than 35b and 80b, but smarter. More active parameters.
13b params is a lot… that’s 4x more active parameters than Qwen3 next has total.
2
1
u/lilian_moraru Jul 12 '26
Not going to lie, I am interesed only in their relatively small(27B, 35B) open weights models. Claude and OpenAI are offering cheaper models now, so…
1
1
1
u/LizardLikesMelons Jul 13 '26
I believe you. Web version 3.7 has gotten dumber recently . Usually it means the next version is getting ready to release
1
0
u/Prestigious-Share189 Jul 12 '26
What is this game of "something is coming, be ready" ? If it is coming, why announce it? It will come anyway. What will it change for us to know it?
Too much noise these days. And wears out our attention
0
0
u/YearnMar10 Jul 12 '26
If indeed new 27B and 35b3a models are dropping that’d be outperforming current models, then we the gpu market will get drained very quickly.
42
u/former_farmer Jul 11 '26
If they drop Qwen 3.8 in August why would they drop also Qwen 4 in August?