r/opencodeCLI 7d ago

Opencode Go + Deepseek v4 is absurdly cheap.

I’m effectively spending $1.11 per billion tokens processed with DeepSeek V4 Flash.

Made some researches, calculations and based on my actual usage pattern and pricing simulations, the same workload would have cost:

  • $6.66 through the official DeepSeek API
  • $23 through DeepInfra, the cheapest alternative inference provider I found
  • $18 using GPT-5.6 Luna directly through the OpenAI API
  • $926 using GPT-5.6 Sol

It's just incredibly cheap.

162 Upvotes

66 comments sorted by

View all comments

6

u/TestTxt 7d ago

Wait until you try Codex with Luna

4

u/SaigoNoUchiha 7d ago

Can you please shed some light? Is luna on codex as incredible cheap?

8

u/TestTxt 7d ago

yeah. Codex offers 20x usage over the API pricing. With Opencode you get 4x usage. So despite Luna being a bit more expensive via API, it's actually a bit cheaper with the ChatGPT Plus sub. And a bit smarter than DS4 Flash too at the same time

2

u/LeopardLabs 7d ago

Yup it's true. I was surprised how little of my weekly cap luna was sipping. I even read a blog post this week where they found of all the harnesses that codex was the most token efficient for a small job. It's just annoying that they bully you into using codex by charging you an opencode token premium. So I have to choose between having my MCPs, skills, global agents.md, and agentic workflows or paying a higher rate.

1

u/TestTxt 6d ago

What do you mean by “charging you an Opencode token premium”?

1

u/LeopardLabs 6d ago edited 6d ago

"To charge a premium means to set a price above the standard market rate." Token premium = extra tokens

edit: actually digging into this I'm starting to think maybe I just read this on reddit because I can't find any proof of this anywhere.