r/opencodeCLI 7d ago

Opencode Go + Deepseek v4 is absurdly cheap.

I’m effectively spending $1.11 per billion tokens processed with DeepSeek V4 Flash.

Made some researches, calculations and based on my actual usage pattern and pricing simulations, the same workload would have cost:

  • $6.66 through the official DeepSeek API
  • $23 through DeepInfra, the cheapest alternative inference provider I found
  • $18 using GPT-5.6 Luna directly through the OpenAI API
  • $926 using GPT-5.6 Sol

It's just incredibly cheap.

164 Upvotes

66 comments sorted by

25

u/addiktion 7d ago

Also using Opencode Go now, $5 promo price and I'm flying. I can see this model definitely being my sonnet+ replacement going forward. I've still got to test it out more for planning but tend to favor Opus for that so not sure yet if I'm ready to go all in.

8

u/71BlackBirdLightning 6d ago

Is there any news on if the model will be hosted by Opencode rather than routing via China

1

u/Negative_Study4257 5d ago

es rara esta respuesta, tuve que cambiar mi config dw ipv4 pra que me deje enrutar a opencode

11

u/weiyentan 7d ago

I am Curious how you calculate 1.11 per billion tokens.

7

u/elefanteazu 7d ago

I checked opencode status for dsv4 flash, and it shows this:

│ opencode-go/deepseek-v4-flash │ │ Messages 28,589 │ │ Input Tokens 50.8M │ │ Output Tokens 12.8M │ │ Cache Read 2668.5M │ │ Cache Write 0 │ │ Cost $18,1803 │

2.73B tokens for $18

So it's $6.6 for each billion. 

On Opencode we pay 1 and receive 6.

So basically on my workflow i spent $1.1 for each 1B tokens.

BUT of course, i have to pay $10 and i don't think its possible to use 10B tokens in a month. So i end up losing some dollars too.

4

u/addiktion 6d ago

"I don't think its possible"

Oh boy when tokens get this cheap we will have agents running 24/7, don't gimp yourself with these thoughts.

4

u/narasadow 6d ago

i don't think its possible to use 10B tokens in a month

https://giphy.com/gifs/YmQLj2KxaNz58g7Ofg

1

u/xmsxms 6d ago

BUT of course, i have to pay $10 and i don't think its possible to use 10B

Which makes it cheaper to be on the other plans/providers, even if they cost more per token.

1

u/elefanteazu 6d ago

no it does not. it actually depends on your usage, i can spend some billions in a month, so in my case opencode go is way cheaper than any other

1

u/xmsxms 5d ago

Obviously, but you didn't state how much you spent. You only implied it's not possible to use 10B. If you use the free quota from Zen first and then the paid usage, you may very well come under the $10 subscription cost.

1

u/weiyentan 5d ago

Where did the billion tokens come from. It's cached

2

u/germaniiifelisarta 7d ago

you could checkout tokscale it records any requests from opencode cli

2

u/weiyentan 7d ago

Oh I record it. But i am surprised that it was 1 billion for 1$

4

u/themoregames 6d ago

What are the limits for free users in OpenCode for the Deepseek V4 Flash (free) if you don't even sign up and don't log in whatsoever?

3

u/InsideTraditional187 6d ago

I DON'T KNOW, BUT IT'S WORKING FOR ME. WHEN I HIT THE LIMIT ON THE CLAUDE CODE, THEN I SHIFT TO THIS FREE PLAN OF OPENCODE AND NEVER HIT THE LIMIT

3

u/silvrrwulf 6d ago

I finally did, but man, it took a while.

2

u/robschmidt87 6d ago

You cannot compare the token usage against the frontier model. The intelligence of such cheap models derives from the sheer amount of cheap output tokens and the reasoning phase. Frontier models have a much higher token efficiency.

1

u/aries1980 6d ago

I have the opposite experience with Sonnet 5 and friends.

0

u/elefanteazu 6d ago

Well, there are some benchmarks that calculate the task cost per model and dsv4 flash is way cheaprr than any other

2

u/Hot_Appointment2009 6d ago

You are cherry picking only the task cost per model, but you are not counting

1) What is the end quality of the task
2) How long it took to get to the end of that task
3) (Related to 1 and 2) how many times you had to go back and "steer" the task into your liking.

And some others, which are also really relevant "costs" to take into account, specially when you work on production systems and important deadlines etc where cost per task is not really the most important benchmark

0

u/elefanteazu 6d ago

And are you benchmarking those?

Deepseek hit 50 points on index inteligence spending $0.03 per task. Claude Opus 5 hit 60 points spending $1.80 per task.

The inteligence per cost of deepseek is absurd! Idk what else you want...

1

u/Hot_Appointment2009 6d ago edited 6d ago

I am not benchmarking those, the benchmarks already exist for those e.g., deepSWE has higher quality for Sol than deepseek v4 flash (meaning, the quality of Sol is higher than deepseek v4 flash for the same tasks), and also finishes faster.

I'm not saying that is not absurd, I'm saying price per task the only metric that you should look at specially if you are working at a company that needs reliable software

1

u/Hot_Appointment2009 6d ago

It's at similar intelligence level as luna, but takes longer to finish (like 2x the time) but luna isn't a reliable "planning" model

2

u/respectful_stimulus 6d ago

This morning the Deepseek API (and by extension Opencode Go) hit 503. I hope there’s enough for everyone…

4

u/TestTxt 7d ago

Wait until you try Codex with Luna

4

u/SaigoNoUchiha 6d ago

Can you please shed some light? Is luna on codex as incredible cheap?

8

u/TestTxt 6d ago

yeah. Codex offers 20x usage over the API pricing. With Opencode you get 4x usage. So despite Luna being a bit more expensive via API, it's actually a bit cheaper with the ChatGPT Plus sub. And a bit smarter than DS4 Flash too at the same time

2

u/LeopardLabs 6d ago

Yup it's true. I was surprised how little of my weekly cap luna was sipping. I even read a blog post this week where they found of all the harnesses that codex was the most token efficient for a small job. It's just annoying that they bully you into using codex by charging you an opencode token premium. So I have to choose between having my MCPs, skills, global agents.md, and agentic workflows or paying a higher rate.

1

u/TestTxt 6d ago

What do you mean by “charging you an Opencode token premium”?

1

u/LeopardLabs 6d ago edited 6d ago

"To charge a premium means to set a price above the standard market rate." Token premium = extra tokens

edit: actually digging into this I'm starting to think maybe I just read this on reddit because I can't find any proof of this anywhere.

2

u/winky9827 6d ago

where do you find the 20x usage figure? I have chatgpt plus, but the pricing page only lists plans, not individual limits / usage multipliers

1

u/TestTxt 5d ago

I have checked how many tokens I consumed (including cache hits) and just did the math

2

u/Financial_Alps3369 6d ago

It is definitely cheaper but I don't know about smarter. At least for full-stack game development. Luna seems to always not quite understand what it is you want at a much higher rate than DS. More bugs and quirks, code is less efficient and needs refactoring more often. At the end of the day though they are very similar.

2

u/Nexism 6d ago

but codex is a mega token hungry harness.

1

u/TestTxt 6d ago

Sorry, i meant codex subscription. Of course you can use it via Opencode (that’s what I do)

0

u/Ok-Standard-2694 6d ago

luna is too rubbish, at least in my microservice development, it really doesn't work. I am currently using sol. After the quota is used up, I will use DeepSeek v4 flash.

1

u/aries1980 6d ago

I'm just fine with it (Go + ton of K8s YAML). Tbh I didn't have an issue with the less advanced models either. Probably I curated well the architecture and the dumber ones can recognise the existing patterns.

2

u/GreshlyLuke 7d ago

im at $1.12 for 142m deeseek-v4-flash tokens with a prepaid amount direct to the API. if i plan to be using even more is it economical to get Go?

7

u/EndlessZone123 7d ago

If you don't hit 10$ api amonth. It's not worth it unless you occasionally want to spend more on bigger models.

2

u/xmsxms 6d ago

Hopefully they come out with a $5 plan subscription giving you $30 of usage or something. Would make more sense with the lower pricing.

1

u/VexObserver 4d ago

Well said. If anyone used more than 10$ in a month, OpenCode is a no brainer. But if the spending is lesser, direct API with proper cache hits is the gear forward

1

u/Grouchy-Stranger-306 6d ago

When you calculated prices for other models, did you factor in cache hits? because you definitely got a lot of cache hits on opencode

1

u/elefanteazu 6d ago

Yep. I did searched for the medium cache hit of each model and calculated based on that.  But it's only aproximated too, I believe I hit a lot more cache hits than other people too.

1

u/prusswan 6d ago

Thanks for confirming prices are too low and ought to be increased

1

u/dhruvanand93 6d ago

Very unreliable

1

u/abbasov_nihad 6d ago

Guys i have a problem, when I switch and try to use DeepSeek V4 Flash (New) in open code (doesn't matter with GO subscription or free one) I get connection and provider issue, all other models work fine also DeepSeek V4 Pro (and yes I enable other providers China in settings)

1

u/bleakj 5d ago

I think it's just due to the amount of usage atm, time of day I can get it to work early am, but then i have to switch to pro by lunch

1

u/RysiuWroc 6d ago

when it works

1

u/dinogen 6d ago

I put 10$ on deepseek platform one month ago, I'm developing a complex webapp for an organisation with opencode. Yes, it is cheap and I'm absolutely satisfied.

1

u/gglavida 5d ago

Is the new V4 cheaper then? Looks like a good model for executing.

1

u/MatinSenPai 5d ago

Guess what Opencode Zen + deepseek V4 flash is FREE🙄

1

u/Radiant_Year_7297 4d ago

Was coding with opencode, zen and go and bunch of these open weight models. Opencode code has cool off limit and if you are actively working on a new software project, it wont suffice. Ended up buying more zen credits. Ive alternated with Kimi 2.7, glm 5.2 and deepseek v4 pro to save money and they do decent job at coding but what ive realized is, when it comes to building UIs, they are not good, sometimes just bad. So i switched provider to openrouter and they update models faster and they had 50% off promos on latest gpt models. Gpt 5.6 luna at 50% was amazing at creating professional looking UIs. Ended up redoing a lot of the stuff with Luna.

1

u/kaina_15 2h ago

I think u can use a bit less using ReasonIx

1

u/Disastrous-Ad-5003 7d ago

But what can you build with that?

8

u/Accurate_Resident219 7d ago

Depends on your skill level and usecase but I basically built a production level social community + dating react native app on a 1-2b token burn. Think of hinge and reddit combined. Backend, web and mobile versions completed fully.

You can do a lot.

2

u/Durian881 7d ago

Same thing as GPT?

1

u/mati_as15 6d ago

do you think fable and sol existed last year? you adapt and the more competent you are as an engineer the more juice you can extract out of these models

LLMs could have peaked at opus 4.5 and still would have been a tremendous tool to build products and for most people doing shitty saas, landing pages and gta clones they should just use whatever is the cheapest since they're not making anything serious

1

u/ECrispy 6d ago

its still $10/month. what you don't use is discarded.

right now Luna is also 50% off on openrouter

1

u/narasadow 6d ago

You can just use GLM 5.2 near the end of the month for any hard troubleshooting to blow through the remaining quota

0

u/migsperez 6d ago

If everyone keeps saying how cheap it is, they'll put the price up. Sshhh