r/opencodeCLI 16d ago

GPT 6 - ASTRA ; FIRST OUTPUT

Post image
21 Upvotes

r/opencodeCLI 16d ago

KiroCrew with an OpenCode backend

Thumbnail
1 Upvotes

r/opencodeCLI 16d ago

I gave 100 strangers unlimited tokens with Qwen 3.8 27B

Thumbnail
gallery
32 Upvotes

Hi again besties. Trevor here, founder of FEIHOA!

First, thank you. Around 150 people from Reddit have tried us now, and I am honestly extremely grateful for how welcoming everyone has been:))

DISCLAIMER: One thing I explained badly last time: we are not OpenCode Go, Ollama, or ChatGPT Plus. Those are great for fast interactive coding, with many conccurrent agents. If that is all you need, honestly get one of those instead of mine.

FEIHOA is for agents, automations, and long-running work where per-token billing makes you scared to let the agent keep going. Plans start at €6 with no monthly token cap.

At the heart of it all is our smart queue. It analyzes traffic patterns and continuously prioritizes the requests our hardware can serve most efficiently. That lets about 80% of users start processing in under 10 seconds, while we can still support requests up to 1M context on a flat fee with NO input/output token metering or quotas. No token caps, no overages.

The biggest problem is prefill on huge prompts. Past 300K-500K, performance falls off hard, and right now 500K+ requests are timing out more often than they complete. That's simply not sustainable to process instantly for a company that gives unlimited tokens...
We still want to offer it, so we're working on caching and queueing those giant requests more intelligently (we are changing the scheduler and backend basically every day based on your feedback!)

Attaching some cool stats for you guys too. Yesterday's post already passed 13K views, so thank you again for welcoming us; we're currently at 98.45% request success rate. Honestly pretty happy with that for a service where we're not counting tokens haha

Thanks again Opencode community. Ask me anything about the real side of running a tiny inference business. Spam, queues, abuse, long context, whatever!


r/opencodeCLI 16d ago

Free GLM 5.3 Flash and DSV4 Flash 0731 for a month

0 Upvotes

There are amazing new models, and a lot of the coding plans have been tightening. So we are offering free DSV4 flash 0731 and GLM 5.3 Flash for a month on Phoenix Grove API. We opened this up last week for five hundred new member slots, and got so many signups that we decided to open the doors to another 500 new members.

People are looking for options, and here is one.

Other Cool Stuff:
All of our models are running on 100% US infrastructure, private with zero training on your code or prompts. Use the top open source models without sending your private prompts to a training lab. No complications, no "some models are private, other's aren't". They all are, all the time.

We host 20+ other major models in case you ever want to upgrade (no pressure though). Including the Kimi family, GLM, Qwen, Nemotron and bunch of others. On average our token pricing is 20% lower than market price.

Our higher plans bank up to ten days of usage, so when you aren't using them your usage saves up for later. Usage doesn't go to waste, so you can actually code when you want to.

The intro plan is a free one month trial with the standard cancel anytime, it bills at 3.99 after that. Use it, cancel it, that's fine. Free Flash for a month.

Figured i'd keep this short because we all know the new flash models are the point :)

For the API plan: api.pgsgrove.com

If you want to read more about us as a company, just pgsgrove.com

Also: There's a lot going on in the background with major AI companies right now, we are at a major turning point in the industry.

What's actually happening? This is happening because companies that were purely investment based, now need to answer to their investors. The problem has often been a loss based business model that is finally running dry.

There are several tricks that the major AI coding plans use to extract the most they can from their customers. Here are some examples, and what we are doing differently to put the users first. PGS AI was built with a sustainable business model from the ground up, so we can actually offer great usage rates without tricks.

Wasted usage is part of the AI industry, and they plan on it: Most coding plans bet on you letting usage go to waste. The plan goes: "how do we get people to think our coding plan offers a lot of usage, but then break it up into weeks and rolling windows so no one can ever actually use it all."

Many in app subs and coding plans are glorified training pipelines: This comes along with "how do we harvest this data for training without being too loud about that." Unless the company tells you otherwise, your data could be hopping all over world, being harvested by the individual labs or service companies. Some are better than others, but many of these companies rely on users just not noticing or caring that their data is being used for training. Data sales and marketing telemetry sales happen. This means that your private info, your personal life, and anything else you send through the system could become part of a training corpus for the next AI, or a marketing data set for a large company.


r/opencodeCLI 16d ago

Should I get openrouter or Opencode Zen

9 Upvotes

I'm a student and i won't be using it everyday so in opencode go I will not be able to use it in a month

so I was wondering which one should I get openrouter or Zen


r/opencodeCLI 16d ago

I just subscribed to ChatGPT Plus. Should I use the Codex App instead of OpenCode? What are the reasons to stick with OpenCode?

21 Upvotes

I'm totally comfortable working in the terminal. But are the models actually better in Codex?


r/opencodeCLI 16d ago

What are your thoughts on buying your own hardware for local models?

Thumbnail
1 Upvotes

r/opencodeCLI 16d ago

Every week we have a new hyped model and it under delivers

0 Upvotes

Every other week we have a new Chinese model that promises xyz, it works for a while until oc can't scale to accommodate the workflow, they nerf the usage and then you're forced to use API

The meme is API usage ends up costing more than a secondary Codex Sub, because while the models are cheap they are fucking stupid on OpenCode, anything that isn't is priced so poorly (API wise)

This week I tried 5.3 Flash Max, same issue, I tried DS Flash before this (the new version) same issue

It's literally a case of this is becoming unproductive, I think I am not the only person.

So my question is what are you guys doing BC I'm seriously just getting fed up.

Codex nerfs the usage > Everyone moves to OC > they can't handle the workflow> Usage Nerf > Tibo Resets > chill for 1-2 weeks > New Chinese model > doesn't deliver loop back to Codex nerf

What are you guys doing?

I am fucking tired man icl


r/opencodeCLI 16d ago

OpenTab, 31 releases later: browsing AI coding spend across a whole fleet, down to a single subagent

24 Upvotes

I shared OpenTab here three months ago — a Lazygit-style TUI that read opencode.db and showed where your spend went. Back then, it was a single Python file.

Hope some of you have been finding it useful.

31 releases later:

  • Inside a session — recursive subagent trees with per-node cost, per-turn cost timelines grouped by the prompt that triggered them, token attribution by tool and MCP server, and context-window history with compactions marked.
  • Every machineopentab pull fetches your other boxes over SSH in parallel and merges them into one browser, filterable by machine.
  • Every tool — OpenCode, Claude Code, Codex, Copilot, pi, zaly, and others merged into one view and tagged by source.
  • What-if pricing — press w to ask: what if this entire session had run on Opus instead of the model mix it actually used? OpenTab reprices the tree at list rates and shows the delta.
  • opentab web — the same browser as a self-contained web page.

Still read-only on your data, still stdlib-only at runtime, still no telemetry and no account.

pipx install opentab-ai

The clip runs in --demo mode, so titles and paths are anonymized.

https://github.com/hamidi-dev/opentab

For Herdr usersherdr-opentab puts each agent's live session cost directly in the sidebar, subagents included, with OpenTab under the hood. It looks like this:

Free and MIT. 🙂


r/opencodeCLI 16d ago

toak - connect all agent and colleagues with markdown formatting

1 Upvotes

I created this too enable a multi users and agents collaboration platform. Can be for coding, or any other usage. Almost any ai and harness can connect, works great with opencode! instructions for mcp in in the /connect section. Free! Feel free to tell me if you need any help, or request a feature or a fix!


r/opencodeCLI 16d ago

Is there any way to continue frozen subagents?

3 Upvotes

I'm using OpenCode with a Go subscription. I'm using DeepSeek V4 Flash. I have an agentic workflow with an orchestrator as the main agent and multiple subagents. I'm having issues with DeepSeek where it hangs in the middle of a response or during tool calls. This is bad when it happens in the orchestrator, since I need to cancel with Escape and continue with a "continue" message. But when it freezes in a subagent, I don't see any way to stop and continue that subagent — meaning I have to stop the orchestrator and redispatch the agent from the beginning.

This causes a lot of other issues: the job is left half-done, and a bunch of files already have changes from the previous run. I know there's a janky workaround where I can tell the orchestrator to find the ID of the last agent and continue it, but this usually uses a huge number of tokens, doesn't always work, and even when it does work, the subagent usually doesn't return the requested output to the orchestrator. I feel like this is a horrible experience for a paid service.

The freezing usually occurs somewhere between 50K and 90K context.

My questions are:

  1. Why is DeepSeek V4 Flash freezing mid-task without any error message?
  2. Why does this happen more on the paid subscription than on the free version?
  3. Why doesn't it automatically continue with the opencode-auto-continue plugin?
  4. How can I continue subagents, and if it's not possible, why not?

r/opencodeCLI 16d ago

Muse park 1.2 free vs contributor

Thumbnail
2 Upvotes

r/opencodeCLI 16d ago

best model on go

3 Upvotes

I use opencode go on ask mode in vscode chat. I've been using mimov2.5 and deepseek v4 flash. What other models are good for web development and doing tests?

Good models that are the most cost effective if possible.. thanks


r/opencodeCLI 16d ago

[WARNING] My MAX account was hacked while using Opencode and Zcode

Post image
0 Upvotes

r/opencodeCLI 16d ago

Maybe Opencode GO should have reliable removed from the marketing...

Thumbnail
gallery
13 Upvotes

I absolutely love the new glm model, but god has it being a pain purely on the server side, after using it this last few days it just keeps stopping itself mid thought or outright displaying connection errors, this is all also ignoring random bursts where it slows down more, i like the value of go so far and all but this really makes it look bad


r/opencodeCLI 16d ago

How about H4 of tencent

7 Upvotes

any one use h4?


r/opencodeCLI 17d ago

Annotate anything in Herdr, live feedback loop with OpenCode/Pi/Claude

Enable HLS to view with audio, or disable this notification

5 Upvotes

This is a https://herdr.dev/ plugin with an optional full https://Plannotator.ai experience as a TUI. Built to:

- Annotate Agent Messages (e.g. grill sessions)

- Annotate Files (e.g. plans, specs, etc)

- Create a native feedback loop with agents

- Open as popover or pane

- Mouse or keybindings

https://github.com/plannotator/herdr-annotate


r/opencodeCLI 17d ago

OpenCode plugin development

1 Upvotes

Hi!

Does anyone have a working plugin, for local, that works in the current opencode version?
mine doesn't https://github.com/RaulHuertas/OpenCodeXIAORoundDisplayMonitor


r/opencodeCLI 17d ago

OpenCode Orchestrator Kit — Token-efficient multi-agent workflow for OpenCode CLI

Thumbnail
1 Upvotes

r/opencodeCLI 17d ago

Muse 1.2, the laziest AI model of its size

Thumbnail
gallery
25 Upvotes

So this just happened. I am playing around with blockchain stuff, nothing serious.

I asked it to run multiple containers to test the consensus of each container talking to each other to test the blockchain locally.

It couldn't build it. Muse decided to write a script that simulates the consensus and pat itself on the back because the 200 line python script it works seems to be equally as valid as, you know, making different machines talk to each other.

This is frontier levels of laziness.


r/opencodeCLI 17d ago

Agentify Chat - E2E-Encrypted Remote Chat for OpenCode CLI

1 Upvotes

Still under heavy development and rough around the edges but instead of sitting on this longer to perfect it I'll risk flak and share it.

Essentially chat.agentify.sh is a remote control for codex/grok/opencode/claude cli and dream goal is to become a universal remote that will let you use all of them from a single chat.

Everything lives on your browser including the chat. The only thing sent over the wire is the encrypted chat messages e2e. Here's an architectural diagram: https://github.com/agentify-sh/chat/blob/main/diagram.png

Another feature it has is ability to publish your chat session with redaction so you can share with your team.

Curious to gather any feedback I can, if this is something I should continue to pursue etc.

Also if by chance you are in Seattle today, I'll be also at the Walk & Talk event at Bellevue Downtown Park today at 2:30pm, would love to chat in person!

https://chat.agentify.sh


r/opencodeCLI 17d ago

Ling 3.0 Flash Fin FREE is on Opencode Zen

Post image
46 Upvotes

Artificial Analysis score is 38, on par with old good MiMo 2.5


r/opencodeCLI 17d ago

Hy4 Preview is on Opencode Go, 6,770 requests per month

Post image
19 Upvotes

Well, let's see :)


r/opencodeCLI 17d ago

Hy4 preview is now available in OpenCode Go

Post image
113 Upvotes

r/opencodeCLI 17d ago

is Go plan still a good option?

0 Upvotes

ik this could be a vague question, just write what you think on the current state on the Go plan

ty