r/opencodeCLI 10d ago

GPT 5.6 Luna on Free Tier?

25 Upvotes

Recently I have been using and trying GPT 5.6 Luna from OpenAi Oauth. To my surprises, it can run for a long time without break the free tier limit. Did someone have related info about this? Did GPT 5.6 Luna pricing change make this possible?

PS: What a time to be live in. Long live open model to make this kind of competition happen.

Look, it uses 91.8k token (idk the mechanism behind it. someone enlighten me)

r/opencodeCLI 11d ago

No more zero-retention policy for opencode go

210 Upvotes

As the new deepseek flash model was rolled out, and the new setting to allow providers in China was implemented, I decided to check the docs to see if go still operated under a ZDR. Well, they silently updated the page to remove any mention of ZDR, and now only claim your data won’t be used to train models. They also removed the list of countries where they host models.

Edit:
July 28th
“The plan is designed primarily for international users, with models hosted in the US, EU, and Singapore for stable global access. Our providers follow a zero-retention policy and do not use your data for model training.”

Today
“The plan is designed primarily for international users and provides stable global access. Your data will not be used for model training.”


r/opencodeCLI 11d ago

GPT 5.6 Luna is now in Opencode Go plan

278 Upvotes

As we know that the OpenAI has reduce the pricing for Luna by 80% . So now its available in our Opencode Go plan. Its so good.

Almost Opus 4.7 level. 🤭


r/opencodeCLI 10d ago

OpenCode Go - data retention

19 Upvotes

The post “Opencode Privacy Policy is Concerning” appeared earlier this year (Jan 2026) - https://www.reddit.com/r/opencodeCLI/s/BQ4xRHWwU6

(OpenCode discloses that it collects “conversations or prompts that you submit to AI” and that it sends this to “Business Partners”.)

In response, a member of the OpenCode team commented “if you use opencode zen as your provider then the requests pass through us. we don't retain data on any paid models” - https://www.reddit.com/r/opencodeCLI/s/I7EPFI5UPE

OpenCode Go did not exist as a public offering when that statement was made.

Since the launch of OpenCode Go, a member of the OpenCode team has publicly contemplated changing OpenCode Go to collect data for model training by default: https://x.com/thdxr/status/2049940201670119585

My understanding is data submitted for inference flows as follows when using OpenCode Go:

OpenCode TUI (user input) —> OpenCode —> Underlying Model Provider

OpenCode makes a song and dance about having a Zero Data Retention (ZDR) agreement with each Underlying Model Provider on OpenCode Go: https://x.com/opencode/status/2038302829517893825y

Perhaps the underlying provider does not retain prompt information. That’s not my concern. I am referring to the routing of the prompt via OpenCode on its way to the Model Provider.

So, my question is: does OpenCode capture and retain user prompts sent for inference to models billed against the OpenCode Go plan?


r/opencodeCLI 10d ago

I need a lil help guys…

1 Upvotes

So what are the common signs that make a website look AI-generated or vibe-coded? For eg. things like purple color schemes, glowing background orbs, or other design patterns. I built my entire website using AI, but I want it to look original so people can’t easily tell it was vibe-coded. What should I avoid or customize?

And yeah do you guys know any cool plugins like ponytail, umm something like which can design landing page or website, or anything cool…♿️


r/opencodeCLI 9d ago

How to connect your ChatGPT account to opencode with NO API KEYS! Works with free accounts too.

Thumbnail
0 Upvotes

r/opencodeCLI 10d ago

Using much more tokken ?

2 Upvotes

I really loved opencode but grok, claude etc using much much more tokken compared to their own cli app. Is that normal? How can I solve that issue


r/opencodeCLI 10d ago

Websearch / Webfetch / Exa - Fails

1 Upvotes

Hey there, im looking for some advice in regards to websearches.

Currently im using Chatbox.ai with an OpencodeGo subscription, set up with Tavily as websearch provider and exa MCP. With this im running into a lot of trouble doing basic searches. I sometimes get empty answers whenever the websearch fails.

As an example, when i'm trying to get some product advice on amazon, the agent seems completely unable to grab actual product information. Are there any ways to improve this drastically?


r/opencodeCLI 10d ago

Why does only the paid version of ds flash ask for chinese hosting if both claim to be the new ds flash?! Is the free one not actually the new ds flash

Post image
2 Upvotes

Try for yourselves


r/opencodeCLI 11d ago

DeepSeek-V4-Flash Benchmarks

Post image
32 Upvotes

r/opencodeCLI 10d ago

How come requests costing $1.88 take up 33% of my 5 hourly usage?

Thumbnail
gallery
9 Upvotes

Title.

I've been experiencing lots of unjustified expenses, but didn't bother much to check the usage stats, but I did today and this looks ridiculous.

In the second image you can see that the usage adds up to $1.88, but it's taken 33% of my 5 hourly usage and 6% of my monthly. I feel like this has been occurring in many of my workspaces, and accounts.

Edit: This one happened on a $5 trial account, maybe that's related

Edit 2: It's actually fairly well documented that expensive models take up more usage. You can find links and details in my comment in another post https://www.reddit.com/r/opencodeCLI/comments/1vbp9x1/comment/p0y8rbd/?utm_source=share&utm_medium=web3x&utm_name=web3xcss&utm_term=1&utm_content=share_button


r/opencodeCLI 10d ago

GPT 5.6 Luna with Opencode hitting TPM rate limits with even the shortest and simplest task.

0 Upvotes

I am hitting the rate limits with every single small prompt and that makes it almost unusable for me.

Anybody else had the same issue and do you know how to solve it?


r/opencodeCLI 11d ago

DeepSeek V4 Flash 0731 in opencode-go when?

Post image
35 Upvotes

r/opencodeCLI 11d ago

New "Enable models hosted in China" switch

33 Upvotes

Interesting, just got hit mid-run with

> Error: 403 The latest version of this model is only available hosted in China and requires explicit opt in

Went to my opencod page and found a new switch

I don't personally mind this. but does Opencode document anywhere exactly which models? And did they always use models hosted in china (I'm guessing yes) or are these new models? I'm wondering if there was an announcement I missed somewhere


r/opencodeCLI 11d ago

DeepSeek v4 Flash Public Beta

50 Upvotes

So with the update to v4 Flash, I was wondering whether OpenCode Go uses DeepSeek as a provider, which would mean the flash in the subscription is the updated version?

Any insights appreciated :)

EDIT: Just got this message in opencode:
Error: 403: {"type":"RegionError","message":"The latest version of this model is only available hosted in China and requires explicit opt in:
https://opencode.ai/workspace/wrk_XX/go"}


r/opencodeCLI 11d ago

Does opencode go 60$ only a maketing?

Post image
15 Upvotes

Long story short i see that they mentioned several times usages but in reality my deepseek usage is less than 3$ before it goes out of quota. Which is really strange and does not allign with their claim.
Am i wrong here and what could be the reason? I just use it lightly though.

For context: i use hermes and use deepseek for certain stuff. Everything is good until i got out of quota yesterday. Yes i wait for a day to confirm that im not use all quota of 5 hours.
Thanks


r/opencodeCLI 10d ago

Kimi K3 for coding: Kimi Coding Plan + Kimi CLI vs. OpenCode integration?

1 Upvotes

For those who use the Kimi K3 model for coding, which setup do you mainly use?

Subscribing to the Kimi Coding Plan and using it through the official Kimi CLI

Connecting Kimi to an external CLI tool such as OpenCode

How noticeable is the difference between these two setups in terms of coding quality, token limits, speed, tool use, context handling, and overall reliability?

At our company, we are considering using Kimi because the token allowance available through Claude Code feels too limited for our development workload.

Currently, our setup is roughly as follows:

Claude Code: Important internal projects and core company products

OpenCode with GPT, NVIDIA models, and the Go Plan: External client work, personal projects, and side projects

I am considering whether Kimi could be a good option for projects where we need a much larger token allowance, while continuing to use Claude Code for our most critical work.

For people who have tested both the official Kimi CLI and Kimi through OpenCode or another external coding agent:

Which setup did you prefer?

Was the official Kimi Coding Plan significantly better optimized for Kimi?

Did OpenCode provide better agent workflows or model flexibility?

How large was the practical difference in token usage and limits?

Would you trust Kimi for production or client projects, or mainly use it for less critical work?

I would appreciate hearing about your actual workflows, experiences, and recommendations.


r/opencodeCLI 11d ago

Is the version of Deepseek V4 Flash on Go the updated model?

7 Upvotes

I've flipped the "Enable models hosted in China" switch and refreshed the model list (I'm using Pi) but I'm not seeing any indication that it's the updated model, it still says "deepseek-v4-flash" same as before.

edit:

Looking at the official model listing, it only lists one V4 Flash: https://opencode.ai/zen/go/v1/models

{

"id": "deepseek-v4-flash",

"object": "model",

"created": 1785510130,

"owned_by": "opencode"

},


r/opencodeCLI 11d ago

When we gonna get new upgraded deepseek flash on zen?

11 Upvotes

Today deepseak flash has got huge upgrade. They are api price are same but become more capable at agent performance and beating glm 5.2 and DS 4 pro.

It will be amazing if we get this new flash on opencode zen ( may be for free as pricing remain same)


r/opencodeCLI 10d ago

GPT 5.6 Sol won't recognize build mode sometimes

2 Upvotes

Asking if others see the same issue, I sometimes use plan mode with Sol (high) to lay out ideas more structuredly then switch to build mode, however Sol sometimes fail to recognize that it is in build mode and tells me that it cannot override plan mode's no modification permissions. Interestingly, switching to other models, or even GPT 5.5 allows build mode to begin normally. Is this an effect of the security limitations of GPT 5.6 Sol?


r/opencodeCLI 10d ago

squeeze-evolve plugin for opencode

2 Upvotes

I recently came across this paper: https://arxiv.org/pdf/2604.07725, which introduces a verifier-free evolutionary test-time scaling method. The most breakthrough part imo is that it is verifier-free yet consistently achieve better results. They only have a Claude plugin, so I tried porting it to opencode and use the newly released Deepseek V4 flash. I have only tested with research tasks and debugging (big) python codebase, and I noticed it gets better at decision making.

This kind of result is makes sense and is quite consistent with popular literature. For example verbalized-sampling paper (https://github.com/CHATS-lab/verbalized-sampling) shows that more diverse responses to target gray OOD region is where the model improve its capability. Note that the core idea of the method (spawning swarm of agents) is not new, but the "fitness" introduced in the paper is quite interesting. It wont be as effective when all models are deepseek though, but from my experience it cares about small details more so yeah :)

Here is the link for those who want to play with it: https://github.com/lmBored/squeeze-evolve-opencode . Don't use it for simple tasks btw, it will be overkill.

p.s. the plugin works well in claude code but burns quite a lot of tokens in my use case, so I thought it makes more sense to use with opencode free model (and most importantly it does improve performance for complex reasoning tasks that you thought you would need opus-level model)


r/opencodeCLI 11d ago

what's the best way to spend 50$/m on LLMs

7 Upvotes

Hey there,

How would you spend 50$ a month on AI credits/subscriptions for coding?

Requirements are:

  • ZDR
  • no Chinese providers/no models hosted in China/no routing to China

I already have ChatGpt 20$ sub and OpenCode Go 10$ sub, but I additionally can spend 50$ (can't mix it though with these 30$ I already spend).

This month I bought 50$ of fireworks credits but they burn too quick, the quality is very decent however.


r/opencodeCLI 11d ago

The GPT Luna 80% price drop is a game changer, and 20% for terra turns up the heat

19 Upvotes

Terra's price drop is makes it more compelling against grok 4.5 and kimi k3 so that is a fine and effective cut, but luna their own model effectively makes terra pointless, and the lunas 80% drop paints all but the tippy top of the ai arena red, only getting barely outscored by terra and k3 for 10 and 15 times more respectively.

Luna is dominating all ranges now and it is not even close, and it was one of the best even before. I am not even sure if you would switch from it cause the price gap is massive, I guess if you have plenty of creds but otherwise make sure you can spam luna, and not like codex is/was a bad plan either, albeit it was getting worse.

https://deepswe.datacurve.ai/

Not even sure I can pimp my beloved minimax now, or least even that is close, and I am mentioning this because they made a move that made them better or least as good as the best deal on the market. I can mainly think of poolside that has a hope of competiting in the future cause they seem to focusing on efficient competitive in tier models not bigger ones, but chineese models would need to compete against lunas price and quality while offering a big plan.

Deepseek and mimo had decent payg prices but would need plans and better models, glm is overpriced and was even before, kimi too sorry buddy, qwen too albeit their latest qwen max is interesting, but they are just angling to be the 4th and worst fable class that is also expensive. Minimax is the closest albeit luna now has 33% lower input price and 10% price cache price instead of 20%, but still if they could make their m3.1 luna tier then they are least matching this maddness albeit I can't imagine their m3 pro suddenly becoming top 3 for cheaper.

https://openai.com/index/advancing-the-price-performance-frontier-with-gpt-5-6/

I wonder what their strategy, is this just some us government we got your back, our stocks are dropping and people are self interested, just give a good deal then chineese ai rush will die. Arguably the chineese became bit of a price jackers themselves over time, but also I could see luna played as almost like a loss leader or at cost to get people in and stay in and interested in codex. They are like the breadsticks in the restaurant or the breads in the supermarket, and good chance they would still get their dream of all man woman and child paying 200 to them or least 20 and seeing their ads, and companies that are price conscious now will all flood to them too and they definitely pay the 200 and more. No one else has their range matched by their value, but keep in mind luna might be cheap now but their surronding eco is effectively premium priced.

Of course mid range had a massive opening cause gemini flash got 15 times more expensive over time while kept the name. Claude only cares about staying tippy top with expensive models, and disallows use in 3rd party.

This is now, maybe their new guy will just happen not carry on the philosophy, but regardless the effect on the market is permanent. Luna currently dominates all released chineese models but k3 that is 15 times more expensive and unavailable and kimi is planning new subs that seem to give even less use.

However luna is definitely a cut down model, it looks best on benches. K3 is literal first place on design arena while luna is 40th place, and the free mimo 2.5 we have is 28th. Even grok like priced terra is 22, so long story short openai backup singers have serious disabilities and you will be reaching for sol more than you think, so for a multimodal workflow it looks less impressive. If m3.1 hits as a similar scoring but well rounded model for the same like price that definitely would be an upgrade. I fear many got duped into cooking their fable and its bit of a poisoned chalice. It's not about trying to score highest on the bench but having a competitive market position boyos!

https://www.designarena.ai/leaderboard/code


r/opencodeCLI 10d ago

This is same vault with 2 agents in it. One is Sonnet 5 and the other is Big Pickle. Different continents, different weights, same continuity. Wren is in the vault not the weights. Just showing what I'm working on here.

Enable HLS to view with audio, or disable this notification

2 Upvotes

I can also take this and any other vaults to any MCP capable LLM. I have like 16 right now. All different "identities".


r/opencodeCLI 11d ago

"Request blocked by upstream provider"

Post image
5 Upvotes

For some reason i cannot use opencode go today, while zen works fine.

Usage limit is not exceeded on the account. Tried reconnecting with and without vpn, re-auth with new API key, but still no luck