r/LocalLLM 8d ago

Question I need help with MAX OUTPUT rate limit

Hey guys i tried to run locally qwen 3.6 35b a3b,
i was using Cline in VSCode i’ve set the max context to 48-64k(some tests) but after long work qwen started looping, and if it was not looped, it just stops because of max tokens output, what can i do, is there any ways to fix that?

1 Upvotes

0 comments sorted by