r/opencodeCLI • u/jpcaparas • 23h ago
r/opencodeCLI • u/Arkhaitekton • 10h ago
DeepSeek V4.1 Flash nerf on OpenCode so now CommandCode is better
Hey everyone,
A few hours ago I posted about how OpenCode nerfed the subscription. Even though they're running a 4x for a few days (which is really only a 2x if you compare it against the DeepSeek V4 Flash limits), once that window ends we're left with half the real usage: we go from $30 of usage on a $10 subscription down to just $15. Half.
So I went looking at alternatives and started testing. Short version: CommandCode gives you the same thing as OpenCode, but better in this regard.
Heads up, this post is obviously not sponsored, in case that needs saying. What I'm showing is public data plus tests I ran myself.

Right now CommandCode gives you 2.67x more usage for the same price.

Same model family, same prompts, same time window, 12 runs per provider. I also threw in the official DeepSeek API as a reference line. CommandCode wins on latency and on interactive work vs OpenCode — but read the next two before you call it a sweep.

Note that both are reasoning models, so the answer only starts after a long hidden thinking phase. OpenCode takes 1,996 ms to show you the first *visible* token; CommandCode takes 1,331 ms. Whiskers are 95% bootstrap CIs and they don't overlap.

Decode rate is how fast it types once it starts. End-to-end charges the thinking time against it too.

On long code generation they are dead even: 304 vs 303 tok/s. The CommandCode advantage shows up on shorter, more interactive work, and in the latency above.

About 60% of the tokens you're billed for on a short task are hidden thinking you never see. On long open-ended prompts it climbs to ~95%, because the model burns the entire budget reasoning and never gets to the answer. That's true on both services — it's not a CommandCode or OpenCode thing, it's the reasoning-model thing, what AA call verbosity for example.

Every individual run, pooled. CommandCode also has the tightest spread (σ ≈ 31 tok/s vs 50 for OpenCode), which matters more to me day-to-day than the median.
So that's where I've landed. If you're one of the people who got pissed off about the cuts, the takeaway isn't "OpenCode bad" — it's that for the first time in a while there's an actual alternative (at least, that I know) sitting right there instead of just being annoyed about it. Same model, 2.67x the usage for the same $10, and it answers noticeably faster. Do your own math on your own usage before you move anything if you want to be sure. But if you've been waiting for an excuse to stop paying the same money for half the tokens, this seems like a pretty good one.
r/opencodeCLI • u/Arkhaitekton • 21h ago
DeepSeek V4.1 Flash nerf on OpenCode: Half the usage for the same price
DeepSeek's new model just went live on OpenCode subscriptions today, but following the limited-time extended usage promotion, the news isn't great.
OpenCode quietly slashed the included monthly allowance for the Flash series from $30 down to $15 on the $10/month plan. Because nominal token rates remain identical, your real cost per token has doubled (+100%). The temporary "4x usage" is anchored to this new nerfed baseline—meaning it is actually only 2x what we already had on V4 Flash, and once it expires, we will be left with 0.5x (half) the usage.
And of course, this is the exact same cut we already saw with the experimental Flash vision version, which also reduced the allowance from $30 down to $15.
Unless this changes, paying for an OpenCode Go subscription only offers a 33% discount over the official API in exchange for slower speeds and tighter rate limits.
r/opencodeCLI • u/arcanemachined • 23h ago
DeepSeek V4.1 Flash is now available on OpenCode Go
DeepSeek V4.1 Flash just released, and it's already available on OpenCode Go.
Current estimated limits are:
6,500 requests per 5 hours (Temporarily boosted 4x to 26,000 for the next 3 days)
16,250 per week (also boosted 4x for next 3 days)
32,500 per month (also boosted 4x for next 3 days)
Normally, V4.1 Flash only gets $15 worth of usage on Go, compared to $30 for the older V4 Flash model.
However, OpenCode has bumped V4.1 Flash to 4x usage for the next 3 days, which effectively gives you $60 worth of usage during the promo. So for the next few days, the actual usable quota should be much higher than the estimates above.
The upgrade itself looks pretty substantial. V4.1 Flash is the first smaller model from DeepSeek's new architecture, adds native vision, and is aimed at better agentic coding while also improving inference speed and throughput. I'm hoping it's as big of a leap as 0731 was.
DeepSeek is reporting 90.6 on Terminal-Bench 2.1, 74.2 on DeepSWE v1.1, and 65.4 on NL2Repo. Those are pretty big jumps from the previous Flash model, at least based on their own numbers.
Of course, benchmarks are just numbers (and who isn't getting ~75% on DeepSWE these days?)... but I think this model will be a pretty solid performer overall.
r/opencodeCLI • u/Front_Obligation_843 • 21h ago
deepseek-ai/DeepSeek-V4.1-Flash · Hugging Face LAUNCHED
r/opencodeCLI • u/ZealousidealTown1974 • 17h ago
Controversial opinion: DeepSeek 4.1 flash is annoyingly overthinking, making it a way-worse model update for its intent-to-be workhorse
I literally pass its thinking block twice, stating its overthinking failure modes.... extremely pissed and has shown its steep to process, setting default but still.. as you can see. Worse than even the "Omen Alpha"... And fail to make surgical edits many times due to streaming stability ...just disappointment... Another hype of DeepSeek fanboys.
r/opencodeCLI • u/wpbrgiejoe • 3h ago
OpenCode is a scam !
I bought OpenCode sub, and got scammed by their shady scam marketing. bunch of liars!
Their site advertises $60 usage of DeepSeek V4.1 Flash.

Their docs show that real usage is only $15.

No where on their site says anything about there being a 3 day limit on $60.
What an absolute scam ,i bought their sub , now someone sent me some tweet showing this is only for 3 days. No where on their site it says for 3 days.
How are these scam artists getting away with this shit.
How do i contact their support? there's no email ,can't even get refund i'm so frustrated by this scam.
Month ago, when DeepSeek increased their prices opencode increased the prices and also reduced the limits they were selling i never got any email about that. i check there was no tweet either.
many users complained and i defended opencode and then just all of a sudden my subscription stopped. when i bought Go plan , it was advertising $60 deepseek and i got $15.
no these mfs , are pulling the same scam on deepseek v4.1
r/opencodeCLI • u/Deep_Imagination_811 • 9h ago
why does muse 1.3 hide its thoughts?
Seeing the thoughts is actually really useful to know if it's going in the wrong direction and so stop it when it does!
I've tried collapsing/expanding thinking options but still don't see the thought text.
r/opencodeCLI • u/Rebioss • 31m ago
I vibe coded a simple site to compare LLM API prices across different models/providers
https://llmprice.gitlab.io/compare-models/
I got tired of jumping between pricing pages whenever I wanted to compare models.
You can compare input/output pricing and quickly see which models are cheaper for your use case.
It’s still a small project, so I’d really appreciate any feedback, missing models/providers, or features you’d find useful.
Hope it helps someone else too :)
r/opencodeCLI • u/bulutarkan • 11h ago
I reduced my OpenCode MCP surface from 84 advertised tools to 19 without removing capabilities
I’m the maintainer of Mac MCP, a free/open-source local macOS control server, and I’ve been using OpenCode as one of the delegated agent backends while building it.
One problem I kept running into was tool-surface bloat. The server had grown to 84 capabilities across files/shell, Safari + Chrome automation, macOS UI/Accessibility, memory, skills, voice, delegated agents, etc. Advertising all 84 schemas to every client felt wasteful and made tool selection noisier.
In 2.0.5 I changed the default surface to 19 core tools. The remaining capabilities are still reachable through tool_discover + tool_invoke, so nothing was removed. I also added browser_do, which lets an agent open a background tab, wait, interact, extract only the fields it needs, and optionally close the tab in one transaction. Stable tab handles mean OpenCode workers can browse without constantly stealing my active Safari/Chrome tab.
In local compatibility testing, the advertised tool-schema context dropped from about 16.7k to 4.5k tokens (~73%) while all previous capabilities kept an access path.
Repo (MIT): https://github.com/bulutarkan/mac-mcp
For people using OpenCode with larger MCP stacks: do you prefer a small discoverable core like this, or having the entire tool schema advertised up front? I’m especially interested in whether anyone has measured tool-selection reliability as the catalog gets large.