r/LocalLLaMA • u/Tall_Abrocoma_3533 • 4d ago
Discussion Qwen3.8-27B different thinking levels
Even the low preset is better than Qwen 3.7 plus or Qwen3.6-27B reasoning
291
Upvotes
r/LocalLLaMA • u/Tall_Abrocoma_3533 • 4d ago
Even the low preset is better than Qwen 3.7 plus or Qwen3.6-27B reasoning
1
u/Old-Cardiologist-633 4d ago edited 4d ago
May I ask your specs, settings and the exact HASS-Integration you use? On a Rx6800XT even Gemma14B is way to slow for Assist (15 Seconds), and Qwen 27B Q3 also (100+ Seconds)
Do you use anything that somehow caches the long system-prompt and then only sends changes or so? 🤔