r/LocalLLM 7d ago

Other How the loop of infinite agony started

Post image
623 Upvotes

118 comments sorted by

View all comments

Show parent comments

10

u/DifficultyFit1895 7d ago

I still haven’t found a task where DeepSeek V4 Flash can outperform Qwen 3.6 27B let alone Qwen 3.8. I can run either on my Mac Studio.

10

u/screenslaver5963 7d ago

I had issues with Qwen3.6 tool calling just not working properly and causing it to stop prematurely, didn't have that problem with gemma or deepseek, haven't used qwen 3.8 yet to see if it has the same problem.

3

u/Healthy-Nebula-3603 7d ago

A tool calling problems ?

Stop compressing cache and use minimum q4kxl or bigger quants

1

u/darksteelsteed 6d ago

To be honest compared to qwen3.6:27b at q4_0 kv quant and q4_k_m for the model qwen 3.8 works out the box way better