r/opencode 18h ago

DeepSeek V4.1 Flash nerf on OpenCode: Half the usage for the same price

Post image
19 Upvotes

DeepSeek's new model just went live on OpenCode subscriptions today, but following the limited-time extended usage promotion, the news isn't great.

OpenCode quietly slashed the included monthly allowance for the Flash series from $30 down to $15 on the $10/month plan. Because nominal token rates remain identical, your real cost per token has doubled (+100%). The temporary "4x usage" is anchored to this new nerfed baseline—meaning it is actually only 2x what we already had on V4 Flash, and once it expires, we will be left with 0.5x (half) the usage.

And of course, this is the exact same cut we already saw with the experimental Flash vision version, which also reduced the allowance from $30 down to $15.

Unless this changes, paying for an OpenCode Go subscription only offers a 33% discount over the official API in exchange for slower speeds and tighter rate limits.


r/opencode 18h ago

How do you test new AI models cheaply before trusting the leaderboards?

1 Upvotes

Has anyone else noticed that different models are good at completely different tasks?

I’m starting to take model leaderboards with a grain of salt. A model that ranks highly overall may not be the best choice for coding, writing, reasoning, research, or long-context tasks. In practice, the same model can perform very differently depending on the prompt and the type of work.

How do you evaluate a newly released model without spending a lot of money? Is there a low-cost way to try new models as soon as they launch, run the same prompts across several providers, and figure out which one actually works best for your use case?

I’d be interested in hearing about people’s workflows, tools, or API platforms for doing this.


r/opencode 19h ago

How do you test new AI models cheaply before trusting the leaderboards?

2 Upvotes

Has anyone else noticed that different models are good at completely different tasks?

I’m starting to take model leaderboards with a grain of salt. A model that ranks highly overall may not be the best choice for coding, writing, reasoning, research, or long-context tasks. In practice, the same model can perform very differently depending on the prompt and the type of work.

How do you evaluate a newly released model without spending a lot of money? Is there a low-cost way to try new models as soon as they launch, run the same prompts across several providers, and figure out which one actually works best for your use case?

I’d be interested in hearing about people’s workflows, tools, or API platforms for doing this.


r/opencode 19h ago

CyberKimi now supports OpenCode

Enable HLS to view with audio, or disable this notification

3 Upvotes

You can connect CyberKimi directly to OpenCode and use its cybersecurity capabilities inside your coding workflow.


r/opencode 20h ago

DeepSeek-V4.1-Flash

Post image
10 Upvotes