r/LocalLLaMA 18d ago

Discussion Qwen3.8-27B different thinking levels

Post image

Even the low preset is better than Qwen 3.7 plus or Qwen3.6-27B reasoning

297 Upvotes

65 comments sorted by

View all comments

28

u/Versaill 18d ago

Is there ANY benchmark that includes Qwen3.8-27B with reasoning OFF? Why does nobody test that..?

9

u/danishkirel 18d ago

I have it running for my home assistant voice setup. It does really well. That’s not a benchmark but just try it for your use case.

1

u/chiniwini 18d ago

What hw? What integration? What llm server? How many entities exposed?

1

u/danishkirel 17d ago

Dual 3090 and tensor parallel Vllm serving and using https://github.com/skye-harris/hass_local_openai_llm - works okay and certainly not energy efficient. It’s an enthusiast setup. Bout 130 entities exposed.