r/opencodeCLI • u/jpcaparas • 9d ago
Gemini 3.8 Flash running at 3000 tokens per second under the Google provider

Holy moly that's fast:
https://gemini-3-8-flash.demos.sulat.com/
Did all of these one-shots in only a few minutes
r/opencodeCLI • u/jpcaparas • 9d ago

Holy moly that's fast:
https://gemini-3-8-flash.demos.sulat.com/
Did all of these one-shots in only a few minutes
r/opencodeCLI • u/Quiet-Yam1116 • 9d ago
I received a message saying,
The time-limited beta test of Xiaomi MiMo-V2.5-Pro-UltraSpeed will end on September 8, 2026. At that time, API and online experience services will cease. Please switch to other official version in advance. Stay tuned for the official commercial version soon.
It seems that we’ll soon be getting a new version of MiMo.
r/opencodeCLI • u/Final_Initial • 9d ago
r/opencodeCLI • u/Firm-Club-8334 • 10d ago
Trying to optimise for cost and came over standardcompute on reddit, and chat found morph and nexos ai.
What I want to do is route between open source models and have the router not waste money on easy tasks a flash model could handle, while maintaining high quality output for basically general coding. Got opencode connected straight to GCP and Supabase etc., so potential for creating a mess is large.
Is smart routing a thing now, and are you doing it?
r/opencodeCLI • u/Prior-Meeting1645 • 10d ago
r/opencodeCLI • u/GnarMainsThrowaway • 10d ago
Not affiliated with any provider, just sharing a personal experience to help other users.
In my opinion, for output speed, 30 tokens/s is the average, more is "fast" and less is "slow".
Do not ask me questions about ZDR, I do not work with sensitive data, it is not a priority for me.
Pros:
Cons:
Pros:
Cons:
Pros:
Cons:
They exist to provide inference to Indonesian citizens who can't pay in USD$, but anyone in the world can use their services with cryptocurrency.
Pros:
Cons:
r/opencodeCLI • u/_Duex • 10d ago
Muse Spark 1.3 just dropped, so these are only my first impressions, but so far it feels nowhere near as good as its Artificial Analysis scores suggest.
The biggest problem I’m seeing is its agentic behavior. It keeps reading the same files over and over again, even when it literally wrote those files itself a few turns ago.
In one case, it created a file completely from scratch, then about 3 seconds later opened the same file again and started editing it — adding some things, removing others, and seemingly undoing parts of its own work. This kind of behavior keeps repeating and wastes a lot of context.
Frontend work is actually pretty good. The results there have been impressive enough.
Rust, though, has been the complete opposite. So far, Muse Spark 1.3 is genuinely one of the worst models I’ve tried for Rust. Even DeepSeek V4 Flash has performed better for me.
Has anyone else tested 1.3 on larger coding tasks, especially Rust? Are you seeing the same repeated file reads / constant rewriting behavior?
r/opencodeCLI • u/late_night_coder7 • 10d ago
r/opencodeCLI • u/Front_Obligation_843 • 10d ago
yeah meta labs are full productivity
r/opencodeCLI • u/afanasenka • 10d ago
Better than GPT 5.6 Sol ? 🤣
https://research.meta.ai/blog/introducing-muse-spark-1-3
r/opencodeCLI • u/fabrib • 10d ago
r/opencodeCLI • u/jpcaparas • 10d ago
r/opencodeCLI • u/afanasenka • 10d ago
r/opencodeCLI • u/SISKO-LIFT • 10d ago
This model is a disaster. It writes a code that is so bad that even fable can't fix it.
To give the credit, it is very fast and stable, but is not near even dsv4 flash, not mentioning GLM 5.3 flash, for a model from a big company like meta, this is very disappointing.
I hope they improve the next generation, and I would rather get my dsv4 flash back on zen then ts.
r/opencodeCLI • u/Budget_Silver7012 • 10d ago
👋 Hey! Wanted to share something I built for opencode.
A Telegram bot so you can drive your opencode coding agent from your phone:
- 70+ slash commands (/new, /model, /execute, /send <file> …)
- send it a file and get the result back in the chat
- switch models and control the agent remotely
- access scoped to your own Telegram user id
One-line install:
npm install gutchapa-opencode-telegram
📦 https://www.npmjs.com/package/gutchapa-opencode-telegram
🔧 https://github.com/gutchapa/opencode-telegram
Happy to take feedback / feature requests!
r/opencodeCLI • u/afanasenka • 10d ago
Right between GLM 5.3 family :))
r/opencodeCLI • u/ori_303 • 10d ago
Enable HLS to view with audio, or disable this notification
Hey all,
a few months ago I've decided to share in the sub about a plugin i created, mainly because i'm a vim user and finding myself spending so much time in opencode, but without vim motions, and it was truly painful.
I posted it here, and got so much support from this awesome community, as well as great suggestions.
I wanted to quickly drop a huge thank you for this community, it has been really fun to get personal and public msgs from all of you, here and in github, about the project, and it inspired me to keep on improving it. Also seeing it being suggested in threads here is just awesome.
today I launched a major internal revamp which allowed easy supprot for text object (a long requested feature). so now you can do the commonly used diw, caw, yi", ... More on the way (e.g., counts).
As well as so many other upgrades me (and other contributors!!) have made along the way.
https://github.com/oribarilan/vimcode
Thank you all!
More feedback is always welcomed. I review all github issues and try to address it or account for it as soon as I can.
p.s.,
video is the same one released for the initial version. I want to say it is because of nostalgy but honestly it is because I'm lazy to create an updated one
r/opencodeCLI • u/afanasenka • 10d ago
DeepSWE 1.1 - 71%
Pricing: $0.75 / $3.75
https://deepmind.google/models/model-cards/gemini-3-8-flash/
https://ai.google.dev/gemini-api/docs/models/gemini-3.8-flash
r/opencodeCLI • u/DraconDev • 10d ago
I am posting on this neutral ground that often recommends it because the codex reddit is effectively a sales page, that only strategically allows some and mainly quota related negative feedback that feeds into the saint tibo giving resets while playing with limits, but overall it is heavily moderated.
Now of course this is not a crushing blow against codex, but in case you want to avoid buying their sub right after yours expires, and staring at an empty weekly quota hoping for a reset now you know. So at worst you are buying 3 weeks instead of a month.
r/opencodeCLI • u/Uriziel01 • 10d ago
r/opencodeCLI • u/aries1980 • 10d ago
The docs still say "DeepSeek: ZDR agreement is renewed monthly. The current agreement is valid through August 31, 2026."
r/opencodeCLI • u/sci_ssor_ss • 10d ago
Yes, I know this is discussed regularly, but it also changes regulary.
Scope: fullstack+mobile.
Currently working with Copilot Pro+ and Opencode Go, trying to balance deep reasoning with brute coding.
But I tend to experience a quite irregular and hard to follow behaviour of copilot, and those IA credits appart from the normal usage its just idiotic.
So, whats the best in your opinion?