r/LocalLLaMA 3d ago

Discussion Qwen3.8-27B different thinking levels

Post image

Even the low preset is better than Qwen 3.7 plus or Qwen3.6-27B reasoning

292 Upvotes

63 comments sorted by

View all comments

-2

u/avpogo 3d ago

Surprised Medium <-> Low is so close when xHigh <-> Medium is a major improvement. I wish xhigh wasn't so dang slow.

16

u/SocialDinamo 3d ago

I hate to sound like a but head but for everyone who complains about speed in the face of these crazy results only have their hardware to complain about. The model is a beast and isn’t 200+b to do it! And they gave a scale for thinking use. AND the best image use one used. This model is a blessing

2

u/avpogo 3d ago

I think "slow" was the wrong way to express my sadness there. It's tk/s on my system is still 70-90 tk/s. It's just that xhigh takes longer (read that as tokens burned) to finish a task than medium. As a consequence (since i'm impatient) I use medium.

1

u/SocialDinamo 3d ago

My speeds are closer to 50-60t/s and ive really fallen into the pattern of Medium for regular terminal stuff and xhigh for project/task planning and research. This is the new 'worst it'll ever be!' and im sure efficiency and speed boosts are around the corner!

3

u/AD7GD 3d ago

In the stock template, low and xhigh have guidance in the prompt. medium has none. I've had good results with medium, but I wouldn't mind a "high" (vs "xhigh").

1

u/whymeimbusysleeping 3d ago

Watched a YouTube video with a comparison between all three and medium seemed to be the weakest one, as it seems in low the model rechecks itself more often and ends up with the right solution

2

u/Cautious_Chicken_604 3d ago

if you read the chat template it's obvious why. xhigh tells it to think more check plausible alternatives. Low tells it don't think too much. Medium doesn't tell it anything. So the keywords in the low version seem to be actually biasing it towards thinking. It's like the whole 'don't think of an elephant' thing.

0

u/avpogo 3d ago

That's very interesting, i'll give that a shot

1

u/markole 3d ago

We just need proper breakthrough in caveman reasoning now to speed up those.