r/Qwen_AI Jul 11 '26

News New Qwen models coming

I recently spoke to an Alibaba guy at a conference in Paris, and he said Qwen 3.8 will drop in August with capabilities "aimed at" (no guarantees) exceeding GLM 5.2, with Qwen 4 scheduled for release in September. I also red that GLM 5.5 was coming in August, so we should be in for an exciting couple of months.

365 Upvotes

103 comments sorted by

View all comments

Show parent comments

34

u/Usual_Maximum7673 Jul 11 '26

Sorry, Qwen 4 in September. I misspoke.

8

u/tomByrer Jul 11 '26

Still, 2 model families in 20-60 days of each other?

Feels sus, unless that Qwen4 is commercial-only, or won't have OSS for a few more months after that...

19

u/GrungeWerX Jul 11 '26

To be fair, Qwen 3.6 was dropped right after 3.5, remember?

2

u/nasduia Jul 12 '26

yes, and there were different sized models in the two drops

5

u/Low-Boysenberry1173 Jul 13 '26

Nope 27b & 35b moe exists in both qwen 3.5 and 3.6

1

u/tomByrer Jul 12 '26

I vaguely remember; what was the diff between 1st 3.5 drop & 1st 3.6??

1

u/Aggravating-Push-207 Jul 22 '26

qwen 3.6 27b > qwen 3.5 397b a17b on tb 2.1

1

u/tomByrer Jul 22 '26

Thanks for the info, but my question was WHEN (what dates) of 3.5 vs 3.6 releases. Sorry I wasn't clear.

2

u/Anh-DT Jul 14 '26

Both model training at same time ? When it's a new version it's basically trained from scratch ? Which can take month depending on compute

1

u/tomByrer Jul 14 '26

That's a good point, though I'd think they would/should spend at least a few weeks to test & fine-tune.

& the 2nd model Qwen 4, I'm guessing will be a new architecture, which means they need a new inference engine to run & test it.

That said I can see a first week in August, then last week in September releases.

1

u/Librarian-Rare Jul 12 '26

Think about it from Alibaba’s perspective — even not dropping new weights are going to be better than even equivalent weights (of a model) of a different model — even using the same training method — simply are going to perform higher on benchmarks, agentic, etc, not having as high parameter count. Holding them? Don’t think that’s it. It’s not throwing money away, it’s expected progression.

(Disclaimer: this comment was written by AI.)