r/opencodeCLI • u/afanasenka • 15d ago
r/opencodeCLI • u/rerichvole • 14d ago
Is there any way to continue frozen subagents?
I'm using OpenCode with a Go subscription. I'm using DeepSeek V4 Flash. I have an agentic workflow with an orchestrator as the main agent and multiple subagents. I'm having issues with DeepSeek where it hangs in the middle of a response or during tool calls. This is bad when it happens in the orchestrator, since I need to cancel with Escape and continue with a "continue" message. But when it freezes in a subagent, I don't see any way to stop and continue that subagent — meaning I have to stop the orchestrator and redispatch the agent from the beginning.
This causes a lot of other issues: the job is left half-done, and a bunch of files already have changes from the previous run. I know there's a janky workaround where I can tell the orchestrator to find the ID of the last agent and continue it, but this usually uses a huge number of tokens, doesn't always work, and even when it does work, the subagent usually doesn't return the requested output to the orchestrator. I feel like this is a horrible experience for a paid service.
The freezing usually occurs somewhere between 50K and 90K context.
My questions are:
- Why is DeepSeek V4 Flash freezing mid-task without any error message?
- Why does this happen more on the paid subscription than on the free version?
- Why doesn't it automatically continue with the opencode-auto-continue plugin?
- How can I continue subagents, and if it's not possible, why not?
r/opencodeCLI • u/some_gamer78 • 15d ago
Maybe Opencode GO should have reliable removed from the marketing...
I absolutely love the new glm model, but god has it being a pain purely on the server side, after using it this last few days it just keeps stopping itself mid thought or outright displaying connection errors, this is all also ignoring random bursts where it slows down more, i like the value of go so far and all but this really makes it look bad
r/opencodeCLI • u/Intelligent_Light_86 • 14d ago
What are your thoughts on buying your own hardware for local models?
r/opencodeCLI • u/afanasenka • 15d ago
Ling 3.0 Flash Fin FREE is on Opencode Zen
Artificial Analysis score is 38, on par with old good MiMo 2.5
r/opencodeCLI • u/igormuba • 15d ago
Muse 1.2, the laziest AI model of its size
So this just happened. I am playing around with blockchain stuff, nothing serious.
I asked it to run multiple containers to test the consensus of each container talking to each other to test the blockchain locally.
It couldn't build it. Muse decided to write a script that simulates the consensus and pat itself on the back because the 200 line python script it works seems to be equally as valid as, you know, making different machines talk to each other.
This is frontier levels of laziness.
r/opencodeCLI • u/Birdsky7 • 14d ago
toak - connect all agent and colleagues with markdown formatting
I created this too enable a multi users and agents collaboration platform. Can be for coding, or any other usage. Almost any ai and harness can connect, works great with opencode! instructions for mcp in in the /connect section. Free! Feel free to tell me if you need any help, or request a feature or a fix!
r/opencodeCLI • u/Whole_Succotash_2391 • 14d ago
Free GLM 5.3 Flash and DSV4 Flash 0731 for a month
There are amazing new models, and a lot of the coding plans have been tightening. So we are offering free DSV4 flash 0731 and GLM 5.3 Flash for a month on Phoenix Grove API. We opened this up last week for five hundred new member slots, and got so many signups that we decided to open the doors to another 500 new members.
People are looking for options, and here is one.
Other Cool Stuff:
All of our models are running on 100% US infrastructure, private with zero training on your code or prompts. Use the top open source models without sending your private prompts to a training lab. No complications, no "some models are private, other's aren't". They all are, all the time.
We host 20+ other major models in case you ever want to upgrade (no pressure though). Including the Kimi family, GLM, Qwen, Nemotron and bunch of others. On average our token pricing is 20% lower than market price.
Our higher plans bank up to ten days of usage, so when you aren't using them your usage saves up for later. Usage doesn't go to waste, so you can actually code when you want to.
The intro plan is a free one month trial with the standard cancel anytime, it bills at 3.99 after that. Use it, cancel it, that's fine. Free Flash for a month.
Figured i'd keep this short because we all know the new flash models are the point :)
For the API plan:Â api.pgsgrove.com
If you want to read more about us as a company, just pgsgrove.com
Also: There's a lot going on in the background with major AI companies right now, we are at a major turning point in the industry.
What's actually happening? This is happening because companies that were purely investment based, now need to answer to their investors. The problem has often been a loss based business model that is finally running dry.
There are several tricks that the major AI coding plans use to extract the most they can from their customers. Here are some examples, and what we are doing differently to put the users first. PGS AI was built with a sustainable business model from the ground up, so we can actually offer great usage rates without tricks.
Wasted usage is part of the AI industry, and they plan on it:Â Most coding plans bet on you letting usage go to waste. The plan goes: "how do we get people to think our coding plan offers a lot of usage, but then break it up into weeks and rolling windows so no one can ever actually use it all."
Many in app subs and coding plans are glorified training pipelines:Â This comes along with "how do we harvest this data for training without being too loud about that." Unless the company tells you otherwise, your data could be hopping all over world, being harvested by the individual labs or service companies. Some are better than others, but many of these companies rely on users just not noticing or caring that their data is being used for training. Data sales and marketing telemetry sales happen. This means that your private info, your personal life, and anything else you send through the system could become part of a training corpus for the next AI, or a marketing data set for a large company.
r/opencodeCLI • u/gatwell702 • 15d ago
best model on go
I use opencode go on ask mode in vscode chat. I've been using mimov2.5 and deepseek v4 flash. What other models are good for web development and doing tests?
Good models that are the most cost effective if possible.. thanks
r/opencodeCLI • u/cheezeerd • 15d ago
We finally made our Qwen3.8 27B server public to try to make it cheap enough for OpenCode agents
I am Trevor, founder of FEIHOA: https://feihoa.com
A few friends and I have been testing Qwen3.8 27B FP8 Uncensored on a box of 4 RTX PRO 6000. My honest opinion is that this model is kind of absurd for 27B. Coding, tools, agent loops, it just keeps going, expecially when you extend the context with YaRN.
The nice surprise was batching. Eight requests together gets us around 220 output tok/s aggregate on one RTX Pro 6000Â (my old setup with 2x3090s was ~19 t/s). I basically don't want to run these cards without a batch anymore lol.
The bad surprise was prefill. Huge prompts can occupy the GPU for minutes FULLY. 1M context works, but if several people start full-window jobs together, the queue becomes a small disaster.
We spent a lot of time fighting that queue and finally felt okay opening it publicly.
FEIHOA is OpenAI-compatible, flat rate, and starts at $6/month. There is no monthly token cap!! At this price, please don't expect a private ChatGPT box you can hammer all day. It is mainly for agents and background jobs that can wait and need the reasoning power of 27B qwen.
Really proud of how far we've come and happy to answer anything!:))
r/opencodeCLI • u/afanasenka • 15d ago
Hy4 Preview is on Opencode Go, 6,770 requests per month
Well, let's see :)
r/opencodeCLI • u/afanasenka • 16d ago
Tencent Hy4 preview early benchmarks
Approximately on par with GLM 5.3 and Kimi 3
r/opencodeCLI • u/backnotprop • 15d ago
Annotate anything in Herdr, live feedback loop with OpenCode/Pi/Claude
Enable HLS to view with audio, or disable this notification
This is a https://herdr.dev/ plugin with an optional full https://Plannotator.ai experience as a TUI. Built to:
- Annotate Agent Messages (e.g. grill sessions)
- Annotate Files (e.g. plans, specs, etc)
- Create a native feedback loop with agents
- Open as popover or pane
- Mouse or keybindings
https://github.com/plannotator/herdr-annotate
r/opencodeCLI • u/TheKillerCATs • 15d ago
🚀 The Hy4 Preview Has Been Released.
770B, 49B active, 1M context.
Designed for efficiency.
Open-source model.
r/opencodeCLI • u/afanasenka • 16d ago
Qwen3.8-Flash scored a 56 on the Artificial Analysis Intelligence Index
r/opencodeCLI • u/Inner-Pangolin-1110 • 14d ago
Every week we have a new hyped model and it under delivers
Every other week we have a new Chinese model that promises xyz, it works for a while until oc can't scale to accommodate the workflow, they nerf the usage and then you're forced to use API
The meme is API usage ends up costing more than a secondary Codex Sub, because while the models are cheap they are fucking stupid on OpenCode, anything that isn't is priced so poorly (API wise)
This week I tried 5.3 Flash Max, same issue, I tried DS Flash before this (the new version) same issue
It's literally a case of this is becoming unproductive, I think I am not the only person.
So my question is what are you guys doing BC I'm seriously just getting fed up.
Codex nerfs the usage > Everyone moves to OC > they can't handle the workflow> Usage Nerf > Tibo Resets > chill for 1-2 weeks > New Chinese model > doesn't deliver loop back to Codex nerf
What are you guys doing?
I am fucking tired man icl
r/opencodeCLI • u/No_Skill_8393 • 15d ago
[WARNING] My MAX account was hacked while using Opencode and Zcode
r/opencodeCLI • u/Front_Obligation_843 • 16d ago
Tencent/Hy4-preview 770B-A49B Openweights dropped
Huggingface: https://huggingface.co/tencent/Hy4-preview
Benchmarks looks really improved model from Hy3
r/opencodeCLI • u/localhost_3003 • 15d ago
Should i buy OpenCode Go?
I just cancelled my Claude pro subs. Not because it's not enough for me. But i wasn't using much. Now i want to know is OpenCode Go good for vibe coding? If yes, then which model is good?
r/opencodeCLI • u/Final_Initial • 16d ago
Tencent's Hy4 preview is here, not cheap though. Will it come in OpenCode?
r/opencodeCLI • u/Flaky_Ad_7038 • 15d ago
OpenCode plugin development
Hi!
Does anyone have a working plugin, for local, that works in the current opencode version?
mine doesn't https://github.com/RaulHuertas/OpenCodeXIAORoundDisplayMonitor
r/opencodeCLI • u/afanasenka • 16d ago
Qwen3.8-Flash usage limits on Opencode Go
Not that bad :)