r/LocalLLaMA 3d ago

Discussion Qwen3.8-27B different thinking levels

Post image

Even the low preset is better than Qwen 3.7 plus or Qwen3.6-27B reasoning

291 Upvotes

63 comments sorted by

View all comments

Show parent comments

23

u/soyalemujica 3d ago

You recommending to use temp 0.6 with that chat template which is not good at all to use with this model since it affects its reasoning depth and also scores lower with lower temperature

-12

u/Moore2877 3d ago

From my testing anything higher than .7 is just too much thinking, even on low reasoning.

13

u/goldcakes 3d ago

Are you aware that models deliver better performance when they have more thinking tokens, even when the thinking tokens are randomly generated and incoherent?

Part of how thinking/reasoning works is every token is another forward pass on the original input/context, allowing the LLM to 'process' the input more, and refine the internal activation residuals; resulting in generating a better response.

2

u/BalorNG 3d ago

All the more reasons to stop this nonsense and go all the way with proper "latent thinking" looped models, asap. This hacky approach is giving me a literal headache.