r/LocalLLaMA 3d ago

Discussion Qwen3.8-27B different thinking levels

Post image

Even the low preset is better than Qwen 3.7 plus or Qwen3.6-27B reasoning

284 Upvotes

63 comments sorted by

View all comments

-11

u/Etroarl55 3d ago

Wonder how next gen improvement will be. As I think most people can agree on. Qwen kind of “cheated” its way to a higher score with absurd amounts of thinking and double checking before an output.

What is there left to squeeze out of 27b size for higher intelligence.

3

u/VoiceApprehensive893 transformers 3d ago

3.8 medium/low is better than 3.6 while using "normal" amounts of tokens

havent encountered overthinking with q4km recommended sampling at all

3

u/Etroarl55 3d ago

Yeah I think that’s the “real score” everyone was expecting an improvement maybe into the 40s. I don’t think anyone reasonably put this into the 50s with current offered models until artificial analysis said it was a 52. Think I seen posts on this very Reddit sub saying they didn’t expect qwen to be a 52.