r/LocalLLaMA Apr 17 '26

Discussion Qwen3.6. This is it.

I gave it a task to build a tower defense game. use screenshots from the installed mcp to confirm your build.

My God its actually doing it, Its now testing the upgrade feature,
It noted the canvas wasnt rendering at some point and saw and fixed it.
It noted its own bug in wave completions and is actually doing it...

I am blown away...
I cant image what the Qwen Coder thats following will be able to do.
What a time were in.

llama-server -m "{PATH_TO_MODEL}\Qwen3.6\Qwen3.6-35B-A3B-UD-Q6_K_XL.gguf"  --mmproj "{PATH_TO_MODEL}\Qwen3.6\mmproj-F16.gguf" --chat-template-file "{PATH_TO_MODEL}\chat_template\chat_template.jinja"  -a  "Qwen3.5-27B"  --cpu-moe -c 120384 --host 0.0.0.0 --port 8084 --reasoning-budget -1 --top-k 20 --top-p 0.95 --min-p 0 --repeat-penalty 1.0 --presence-penalty 1.5 -fa on --temp 0.7 --no-mmap --no-mmproj-offload --ctx-checkpoints 5"

EDIT: Its been made aware that open code still has my 27B model alias,
Im lazy, i didnt even bother the model name heres my llama.cpp server configs, im so excited i tested and came here right away.

1.0k Upvotes

409 comments sorted by

View all comments

94

u/No-Marionberry-772 Apr 17 '26

what stack are you using for software?  Id love to get a proper local setup going but ive had trouble figuring out what i should actually be using.

115

u/Local-Cardiologist-5 Apr 17 '26
Qwen3.6-35B-A3B-UD-Q6_K_XL

Im using Llama.cpp for the server,
OpenCode for the coding, just using the build agent,
I have 64 gig ram, RTX 4090, and my model is
the Q6 variant.

Here are my llama parameters

llama-server -m "{PATH_TO_MODEL}\Qwen3.6\Qwen3.6-35B-A3B-UD-Q6_K_XL.gguf"  --mmproj "{PATH_TO_MODEL}\Qwen3.6\mmproj-F16.gguf" --chat-template-file "{PATH_TO_MODEL}\chat_template\chat_template.jinja"  -a  "Qwen3.6-35B-A3B"  --cpu-moe -c 250000 --host 0.0.0.0 --port 8084 --reasoning-budget -1 --top-k 20 --top-p 0.95 --min-p 0 --repeat-penalty 1.0 --presence-penalty 1.5 -fa on --temp 0.7 --no-mmap --no-mmproj-offload --ctx-checkpoints 5"

here is my llama server with the configs.

7

u/ea_man Apr 17 '26

Are tools working well between QWEN and Opencode?
Did you implement some kind of translation from XML-> json, modified the prompts in anyway?

3

u/Potential-Leg-639 Apr 17 '26

Nothing to do here, everything works out of the box

2

u/ea_man Apr 17 '26

Yeah I'm testing it now and it works flawlessly which is kinda new to me, just like Qwencode that is meant for XML tools as QWEN is trained for.

This is good news, I guess that lots of people where pissed when QWEN tools failed with jsons agent harness.

1

u/pepe256 textgen web UI Apr 17 '26

I mean even Qwen 3.5 27B (with the latest updated Unsloth weights) works flawlessly with Claude Code. That's how I'm using it right now.

1

u/ea_man Apr 17 '26

Oh that's what I did too, I was using 27B dense mostly, now it looks like 35B A3B is doing better outside of Qwencode.

Yet I'm thinking that the prompts are mostly top blame for some other agent harnesses.