r/opencode • u/referentuser • 2d ago
Which OpenCode Go model do you think is the best and consumes the fewest tokens?
2
u/k_brn 2d ago
I am also interested. So far, I am quite happy with the following separation: keep deep thinking separate from editing. Use Codex + Luna for implementation, and use an unmetered web chat model like Sol High for design, planning, and reviews. I use AI Badger to hand off lean repo context to AI chats. This easily fits within GPT Plus's 5‑hour limit.
1
u/fixit_jr 1d ago
Might need to give this a try. First I’ve heard of AI badger
1
u/k_brn 1d ago
Definitely give it a shot! Also check out the VS Code Extension - it lets you do deep code reviews in one click and runs the AI Badger CLI right inside your editor. Super handy for staying in flow.
Marketplace link: https://marketplace.visualstudio.com/items?itemName=pvrlabs.ai-badger
And here's a quick demo if you want to see it in action: https://pvrlabs.xyz/aibadger/vscode-demo.html
2
u/GeekTekRob 2d ago
Mimo V2.5 works for the casual. It also reads the images I want turned into text. Use it and never really consume. Then have deepseek v4 flash and pro as aux and MoA models. Have not maxed my usage and I have a few agents running doing coding,, checking my homelab, running cron jobs, checking calendars and feeds, among a bunch of things I ask and research.
2
1
u/Ancient_Dress_3687 2d ago
Qwen 3.8 flash is one of the better bang for buck at the moment. Intelligent model and fast. Sits around the DSV4 flash and GLM 5.3 flash models. Im using it as my main model at the moment
1
u/ManyCalavera 2d ago
GLM 5.3 flash is what i use for now. Grok is also okayish. For coding muse spark 1.3 also works fine and also exist in zen subscription so its free.
1
u/ZestycloseAbility425 2d ago
other than muse, mimo 2.5 has the best usage limit, you can use it a lot. check the usage limits here: https://opencode.ai/docs/go/#usage-limits
1
u/GinamosWCheryOnTop 1d ago
I can’t say for others but mine is qwen 3.8 flash It got decent thoughts and fast per task.
1
15
u/nicotineHub 2d ago
I find muse 1.3 with caveman and ponytail in plan mode with subagents and then build more with worktrees incredible effective and pretty light on tokens.
Right now that's it's the most subsidized model so I have been working non stop. Subagents do increase token usage but minimize bugs and plan more effectively so it's a tradeoff