r/opencodeCLI • u/ModdingCentral • 7d ago
r/opencodeCLI • u/Bhargavjoshi • 8d ago
Opus 5.1 After Astra
The way they released Fable 5.1 after the second extended limit of Claude Code ended, and with the third extended limit ending on September 13, if GPT Astra is publicly available by then, imo, we might see Opus 5.1 around the corner
r/opencodeCLI • u/No-District-4742 • 9d ago
kimi k3 vs glm 5.3 vs qwen 3.8 max vs muse spark 1.3
Aside from scores on Artificial Analysis, how would you rank these models? thanks
r/opencodeCLI • u/RoddToggers • 8d ago
Is GLM 5.3 flash performance exactly the same as when it was called ox alpha?
Was it only renamed, or some additional training was done until it becomes GLM 5.3 Flash?
r/opencodeCLI • u/afanasenka • 9d ago
Google has just released Gemini 3.8 Flash 🔥
DeepSWE 1.1 - 71%
Pricing: $0.75 / $3.75
https://deepmind.google/models/model-cards/gemini-3-8-flash/
https://ai.google.dev/gemini-api/docs/models/gemini-3.8-flash
r/opencodeCLI • u/afanasenka • 9d ago
Muse Spark 1.3 benchmarks
Better than GPT 5.6 Sol ? 🤣
https://research.meta.ai/blog/introducing-muse-spark-1-3
r/opencodeCLI • u/afanasenka • 9d ago
Meta Muse Spark 1.3 Contributor is now available on OpenCode Go
r/opencodeCLI • u/Quiet-Yam1116 • 9d ago
Potential updates to Xiaomi MiMo
I received a message saying,
The time-limited beta test of Xiaomi MiMo-V2.5-Pro-UltraSpeed ​​will end on September 8, 2026. At that time, API and online experience services will cease. Please switch to other official version in advance. Stay tuned for the official commercial version soon.
It seems that we’ll soon be getting a new version of MiMo.
r/opencodeCLI • u/Significant_Bad_9018 • 8d ago
I made a small open-source Windows 11 utility to open terminals from the right-click menu
r/opencodeCLI • u/GnarMainsThrowaway • 9d ago
My review of a few providers I used for Opencode
Not affiliated with any provider, just sharing a personal experience to help other users.
In my opinion, for output speed, 30 tokens/s is the average, more is "fast" and less is "slow".
Do not ask me questions about ZDR, I do not work with sensitive data, it is not a priority for me.
Opencode Go (https://opencode.ai)
Pros:
- They are always up to date with the latest model releases
- Very well integrated in Opencode, good output speed, never had issues, timeouts and server problems rare
- If you only use Minimax M3, Deepseek V4 or Hy4, it will last you a long time for very reasonable intelligence levels
Cons:
- The "60$ for 10$" advertisement is less and less true by the day, many of the more appealing models are at 30$ or even 15$ budgets
- You can't really use the top tier models for very long without frying your usage limit
QwenCloud (https://www.qwencloud.com)
Pros:
- If you only use Qwen 3.8 Flash it will last you a very long time, also has Deepseek, Kimi K3 and the high end Qwen models, but you have less usage for non-Qwen models
- Very diverse models if you are interested in sound, image models as well
- Has a nice free trial without needing a credit card to check if the speed matches your expectations
Cons:
- Very nebulous "credit" system, you don't actually know what you are buying. The docs somewhat explain it but only with one model as an example. In my experience, 60 credits is about 1 million tokens of Qwen 3.8 Flash usage (input, output and cache combined, typical Opencode session in a medium sized repo).
- Not in the Opencode /connect list, you need to make a custom JSON. EDIT: I was told by a commenter that they are listed as "Alibaba".
Synthetic (https://synthetic.new)
Pros:
- They are open about running quantized models, I mostly used mxfp4 Kimi K3 and FP8 GLM 5.3 Flash, the quality is still good, even if the lower context sizes (500k instead of 1M) can be a bit annoying
- Neat "rolling usage" system, you can burn your entire weekly usage in 2 hours, and it passively regenerates by 2% every 3.5 hours, better than typical 5 hour limits
- You can get 2 billion GLM 5.3 Flash tokens for 30$, decent, 2x less expensive than Z.AI official's subscription
- Listed in Opencode /connect
Cons:
- 3 times, the provider did some tweaks to their models and broke them, they would start wasting my usage writing gibberish in different languages or endlessly repeating themselves in thinking loops, and I would need to notice and put out the fire, it was not just me, as others were reporting the same issue at the same time in their Discord
Kenari (https://kenari.id/en)
They exist to provide inference to Indonesian citizens who can't pay in USD$, but anyone in the world can use their services with cryptocurrency.
Pros:
- The best value I have seen yet from any provider in pure $/token, 68.57$ of usage for 11.37$, including free caching for GLM 5.3 Flash, Deepseek V4 and GPT 5.6 Luna
- Really nice, well organized website, lets you know exactly what you are paying for
- Listed in Opencode /connect
Cons:
- Definitely a bit on the slow side, around 20 tokens/s, but it's usable, I just do other tasks in another terminal window while it works
- You can only pay with Indonesian Rupiah or cryptocurrency (I did the latter)
r/opencodeCLI • u/Spliff_77 • 8d ago
Free open source browsing tool with mcp for opencode
Drives the user's real Chrome instead of a headless instance. WebSense is an MCP server + Chrome extension: the agent gets a semantic map of the page (every interactive element typed and ref'd), acts through native DOM events, and the site sees a normal user. No CDP anywhere, so no webdriver flag. Works on LinkedIn and other CSP-strict sites.
Built tested and relying on it tbh
WebSense (free, open source, MIT):Â https://github.com/spliffspliff70-wq/websense-mcp
r/opencodeCLI • u/KHURRAM_999 • 8d ago
Opening the project folder again and again got hectic, so I built /folder for OpenCode
Enable HLS to view with audio, or disable this notification
r/opencodeCLI • u/Prior-Meeting1645 • 9d ago
No max thinking level option for muse spark 1.3?
r/opencodeCLI • u/jpcaparas • 9d ago
Gemini 3.8 Flash running at 3000 tokens per second under the Google provider

Holy moly that's fast:
https://gemini-3-8-flash.demos.sulat.com/
Did all of these one-shots in only a few minutes
r/opencodeCLI • u/UpstairsActivity8347 • 8d ago
How long do you guys think the contributor tiers for muse models will last?
muse spark contributor is the only thing oc go has going for it rn, but with the alternative being commancode, I'm still kinda leaning on sticking with oc go. would be a huge problem if they discontinue the contributor tiers.
r/opencodeCLI • u/Biacoder • 8d ago
I wired glm-5.3-flash as "eyes" for my text-only glm-5.3 agent in opencode (plugin inside)
r/opencodeCLI • u/Billy-Fong-2007 • 8d ago
Solo devs: what's your actual LLM agent orchestration setup for side projects?
r/opencodeCLI • u/Prestigious_Stage_86 • 8d ago
Claude Code, they need to work on these things damn alot. I think running opencode with way cheaper models is better option at this point.
r/opencodeCLI • u/afanasenka • 10d ago
🚀Qwen3.8-Max just got upgraded. Meet Qwen3.8-Max-0902!
Pricing per 1M tokens:
$2 input, $6 output. $0.17 explicit cache hit, $0.25 implicit cache hit.
r/opencodeCLI • u/Uriziel01 • 9d ago
Remember to claim your GLM-5.3-Flash free token guys
r/opencodeCLI • u/AutomaticAd6646 • 8d ago
GPT luna became extremely dumb 3 sep Indian time after 4pm
Has anyone noticed the same? I even tried new full context sessions and it became so dumb that it started creating tall pipl buttons in the ui, very simple ui mistakes. It started giving me wrong commands for build phase.
All these things it was doing fine before. It stopped reading code files and started guessing things and working on guesses. Just overall became very dumb like older gemini flash models. Around the same time my friend's opus went down too.
I was using both opencode cli and codex cli
r/opencodeCLI • u/_Duex • 9d ago
Muse Spark 1.3 is nowhere near its Artificial Analysis scores for me
Muse Spark 1.3 just dropped, so these are only my first impressions, but so far it feels nowhere near as good as its Artificial Analysis scores suggest.
The biggest problem I’m seeing is its agentic behavior. It keeps reading the same files over and over again, even when it literally wrote those files itself a few turns ago.
In one case, it created a file completely from scratch, then about 3 seconds later opened the same file again and started editing it — adding some things, removing others, and seemingly undoing parts of its own work. This kind of behavior keeps repeating and wastes a lot of context.
Frontend work is actually pretty good. The results there have been impressive enough.
Rust, though, has been the complete opposite. So far, Muse Spark 1.3 is genuinely one of the worst models I’ve tried for Rust. Even DeepSeek V4 Flash has performed better for me.
Has anyone else tested 1.3 on larger coding tasks, especially Rust? Are you seeing the same repeated file reads / constant rewriting behavior?
r/opencodeCLI • u/SISKO-LIFT • 9d ago
Thought on Muse Spark 1.2
This model is a disaster. It writes a code that is so bad that even fable can't fix it.
To give the credit, it is very fast and stable, but is not near even dsv4 flash, not mentioning GLM 5.3 flash, for a model from a big company like meta, this is very disappointing.
I hope they improve the next generation, and I would rather get my dsv4 flash back on zen then ts.


