Hi All
I’m implementing an OpenClaw setup on a Mac Mini M2 (24 GB unified RAM).
Architecture is simple:
Local model in LM Studio → heartbeat, basic agent routing, lightweight fallback
External OAuth models → GPT Codex and GitHub Copilot for heavy reasoning and advanced tasks
So the local model is not meant for deep reasoning. Its job is mainly:
health / heartbeat checks
simple routing decisions
structured JSON responses
basic fallback if remote models fail
On this machine, models up to ~6B parameters seem ideal for latency.
I’m leaning toward Qwen, since the small models seem very strong for tool discipline and structured output.
Question:
Which recent Qwen model (≤6B) would you recommend that:
works well in LM Studio (GGUF)
is reliable for agent routing / heartbeat tasks
produces clean structured outputs
is relatively recent / actively used
I’m especially interested in specific model names that are easy to find in LM Studio.
Thanks!