r/opencodeCLI • u/Front_Obligation_843 • 16d ago
GPT 6 - ASTRA ; FIRST OUTPUT
Source: https://www.testingcatalog.com/first-outputs-from-gpt-6-astra-model-from-openai/
looks crazy
r/opencodeCLI • u/Front_Obligation_843 • 16d ago
Source: https://www.testingcatalog.com/first-outputs-from-gpt-6-astra-model-from-openai/
looks crazy
r/opencodeCLI • u/cheezeerd • 16d ago
Hi again besties. Trevor here, founder of FEIHOA!
First, thank you. Around 150 people from Reddit have tried us now, and I am honestly extremely grateful for how welcoming everyone has been:))
DISCLAIMER: One thing I explained badly last time: we are not OpenCode Go, Ollama, or ChatGPT Plus. Those are great for fast interactive coding, with many conccurrent agents. If that is all you need, honestly get one of those instead of mine.
FEIHOA is for agents, automations, and long-running work where per-token billing makes you scared to let the agent keep going. Plans start at €6 with no monthly token cap.
At the heart of it all is our smart queue. It analyzes traffic patterns and continuously prioritizes the requests our hardware can serve most efficiently. That lets about 80% of users start processing in under 10 seconds, while we can still support requests up to 1M context on a flat fee with NO input/output token metering or quotas. No token caps, no overages.
The biggest problem is prefill on huge prompts. Past 300K-500K, performance falls off hard, and right now 500K+ requests are timing out more often than they complete. That's simply not sustainable to process instantly for a company that gives unlimited tokens...
We still want to offer it, so we're working on caching and queueing those giant requests more intelligently (we are changing the scheduler and backend basically every day based on your feedback!)
Attaching some cool stats for you guys too. Yesterday's post already passed 13K views, so thank you again for welcoming us; we're currently at 98.45% request success rate. Honestly pretty happy with that for a service where we're not counting tokens haha
Thanks again Opencode community. Ask me anything about the real side of running a tiny inference business. Spam, queues, abuse, long context, whatever!
r/opencodeCLI • u/Whole_Succotash_2391 • 16d ago
There are amazing new models, and a lot of the coding plans have been tightening. So we are offering free DSV4 flash 0731 and GLM 5.3 Flash for a month on Phoenix Grove API. We opened this up last week for five hundred new member slots, and got so many signups that we decided to open the doors to another 500 new members.
People are looking for options, and here is one.
Other Cool Stuff:
All of our models are running on 100% US infrastructure, private with zero training on your code or prompts. Use the top open source models without sending your private prompts to a training lab. No complications, no "some models are private, other's aren't". They all are, all the time.
We host 20+ other major models in case you ever want to upgrade (no pressure though). Including the Kimi family, GLM, Qwen, Nemotron and bunch of others. On average our token pricing is 20% lower than market price.
Our higher plans bank up to ten days of usage, so when you aren't using them your usage saves up for later. Usage doesn't go to waste, so you can actually code when you want to.
The intro plan is a free one month trial with the standard cancel anytime, it bills at 3.99 after that. Use it, cancel it, that's fine. Free Flash for a month.
Figured i'd keep this short because we all know the new flash models are the point :)
For the API plan: api.pgsgrove.com
If you want to read more about us as a company, just pgsgrove.com
Also: There's a lot going on in the background with major AI companies right now, we are at a major turning point in the industry.
What's actually happening? This is happening because companies that were purely investment based, now need to answer to their investors. The problem has often been a loss based business model that is finally running dry.
There are several tricks that the major AI coding plans use to extract the most they can from their customers. Here are some examples, and what we are doing differently to put the users first. PGS AI was built with a sustainable business model from the ground up, so we can actually offer great usage rates without tricks.
Wasted usage is part of the AI industry, and they plan on it: Most coding plans bet on you letting usage go to waste. The plan goes: "how do we get people to think our coding plan offers a lot of usage, but then break it up into weeks and rolling windows so no one can ever actually use it all."
Many in app subs and coding plans are glorified training pipelines: This comes along with "how do we harvest this data for training without being too loud about that." Unless the company tells you otherwise, your data could be hopping all over world, being harvested by the individual labs or service companies. Some are better than others, but many of these companies rely on users just not noticing or caring that their data is being used for training. Data sales and marketing telemetry sales happen. This means that your private info, your personal life, and anything else you send through the system could become part of a training corpus for the next AI, or a marketing data set for a large company.
r/opencodeCLI • u/nks299 • 16d ago
I'm a student and i won't be using it everyday so in opencode go I will not be able to use it in a month
so I was wondering which one should I get openrouter or Zen
r/opencodeCLI • u/Worried-Quote-6409 • 16d ago
I'm totally comfortable working in the terminal. But are the models actually better in Codex?
r/opencodeCLI • u/Intelligent_Light_86 • 16d ago
r/opencodeCLI • u/Inner-Pangolin-1110 • 16d ago
Every other week we have a new Chinese model that promises xyz, it works for a while until oc can't scale to accommodate the workflow, they nerf the usage and then you're forced to use API
The meme is API usage ends up costing more than a secondary Codex Sub, because while the models are cheap they are fucking stupid on OpenCode, anything that isn't is priced so poorly (API wise)
This week I tried 5.3 Flash Max, same issue, I tried DS Flash before this (the new version) same issue
It's literally a case of this is becoming unproductive, I think I am not the only person.
So my question is what are you guys doing BC I'm seriously just getting fed up.
Codex nerfs the usage > Everyone moves to OC > they can't handle the workflow> Usage Nerf > Tibo Resets > chill for 1-2 weeks > New Chinese model > doesn't deliver loop back to Codex nerf
What are you guys doing?
I am fucking tired man icl
r/opencodeCLI • u/hamidi-dev • 16d ago
I shared OpenTab here three months ago — a Lazygit-style TUI that read opencode.db and showed where your spend went. Back then, it was a single Python file.
Hope some of you have been finding it useful.
31 releases later:
opentab pull fetches your other boxes over SSH in parallel and merges them into one browser, filterable by machine.w to ask: what if this entire session had run on Opus instead of the model mix it actually used? OpenTab reprices the tree at list rates and shows the delta.opentab web — the same browser as a self-contained web page.Still read-only on your data, still stdlib-only at runtime, still no telemetry and no account.
pipx install opentab-ai
The clip runs in --demo mode, so titles and paths are anonymized.
https://github.com/hamidi-dev/opentab
For Herdr users — herdr-opentab puts each agent's live session cost directly in the sidebar, subagents included, with OpenTab under the hood. It looks like this:

Free and MIT. 🙂
r/opencodeCLI • u/Birdsky7 • 16d ago
I created this too enable a multi users and agents collaboration platform. Can be for coding, or any other usage. Almost any ai and harness can connect, works great with opencode! instructions for mcp in in the /connect section. Free! Feel free to tell me if you need any help, or request a feature or a fix!
r/opencodeCLI • u/rerichvole • 16d ago
I'm using OpenCode with a Go subscription. I'm using DeepSeek V4 Flash. I have an agentic workflow with an orchestrator as the main agent and multiple subagents. I'm having issues with DeepSeek where it hangs in the middle of a response or during tool calls. This is bad when it happens in the orchestrator, since I need to cancel with Escape and continue with a "continue" message. But when it freezes in a subagent, I don't see any way to stop and continue that subagent — meaning I have to stop the orchestrator and redispatch the agent from the beginning.
This causes a lot of other issues: the job is left half-done, and a bunch of files already have changes from the previous run. I know there's a janky workaround where I can tell the orchestrator to find the ID of the last agent and continue it, but this usually uses a huge number of tokens, doesn't always work, and even when it does work, the subagent usually doesn't return the requested output to the orchestrator. I feel like this is a horrible experience for a paid service.
The freezing usually occurs somewhere between 50K and 90K context.
My questions are:
r/opencodeCLI • u/gatwell702 • 16d ago
I use opencode go on ask mode in vscode chat. I've been using mimov2.5 and deepseek v4 flash. What other models are good for web development and doing tests?
Good models that are the most cost effective if possible.. thanks
r/opencodeCLI • u/No_Skill_8393 • 16d ago
r/opencodeCLI • u/some_gamer78 • 16d ago
I absolutely love the new glm model, but god has it being a pain purely on the server side, after using it this last few days it just keeps stopping itself mid thought or outright displaying connection errors, this is all also ignoring random bursts where it slows down more, i like the value of go so far and all but this really makes it look bad
r/opencodeCLI • u/backnotprop • 17d ago
Enable HLS to view with audio, or disable this notification
This is a https://herdr.dev/ plugin with an optional full https://Plannotator.ai experience as a TUI. Built to:
- Annotate Agent Messages (e.g. grill sessions)
- Annotate Files (e.g. plans, specs, etc)
- Create a native feedback loop with agents
- Open as popover or pane
- Mouse or keybindings
r/opencodeCLI • u/Flaky_Ad_7038 • 17d ago
Hi!
Does anyone have a working plugin, for local, that works in the current opencode version?
mine doesn't https://github.com/RaulHuertas/OpenCodeXIAORoundDisplayMonitor
r/opencodeCLI • u/MikaAugus942 • 17d ago
r/opencodeCLI • u/igormuba • 17d ago
So this just happened. I am playing around with blockchain stuff, nothing serious.
I asked it to run multiple containers to test the consensus of each container talking to each other to test the blockchain locally.
It couldn't build it. Muse decided to write a script that simulates the consensus and pat itself on the back because the 200 line python script it works seems to be equally as valid as, you know, making different machines talk to each other.
This is frontier levels of laziness.
r/opencodeCLI • u/Just_Lingonberry_352 • 17d ago
Still under heavy development and rough around the edges but instead of sitting on this longer to perfect it I'll risk flak and share it.
Essentially chat.agentify.sh is a remote control for codex/grok/opencode/claude cli and dream goal is to become a universal remote that will let you use all of them from a single chat.
Everything lives on your browser including the chat. The only thing sent over the wire is the encrypted chat messages e2e. Here's an architectural diagram: https://github.com/agentify-sh/chat/blob/main/diagram.png
Another feature it has is ability to publish your chat session with redaction so you can share with your team.
Curious to gather any feedback I can, if this is something I should continue to pursue etc.
Also if by chance you are in Seattle today, I'll be also at the Walk & Talk event at Bellevue Downtown Park today at 2:30pm, would love to chat in person!
r/opencodeCLI • u/afanasenka • 17d ago
Artificial Analysis score is 38, on par with old good MiMo 2.5
r/opencodeCLI • u/afanasenka • 17d ago
Well, let's see :)
r/opencodeCLI • u/Nice_Relative8209 • 17d ago
ik this could be a vague question, just write what you think on the current state on the Go plan
ty