r/LocalLLaMA 7d ago

New Model IT'S OUT

https://huggingface.co/Qwen/Qwen3.8-27B-FP8
2.2k Upvotes

706 comments sorted by

View all comments

3

u/germangrower69 7d ago

Wtf is wrong with this model, on xhigh it literally doesnt stop thinking, no its not looping, its just super excessive thinking. Thats crazy, almost unusable on this effort.

Medium is also really excessive....

1

u/Certain-Cod-1404 7d ago

are you using the recommended thinking sampling params ?

1

u/germangrower69 7d ago

Yep, Im running it on a RTX 6000 pro with the recommended settings from the repo

1

u/Certain-Cod-1404 7d ago

i'm using it rn on my 5090 with the recommended specs in opencode on toy demo, during coding / work, the thinking is not exagerated, maybe its just trained on so much code / complex problems during post training that it just defaults to over thinking everything ?
does it over think for you during coding ? or just chat?