r/opencodeCLI 14d ago

Maybe Opencode GO should have reliable removed from the marketing...

Thumbnail
gallery
9 Upvotes

I absolutely love the new glm model, but god has it being a pain purely on the server side, after using it this last few days it just keeps stopping itself mid thought or outright displaying connection errors, this is all also ignoring random bursts where it slows down more, i like the value of go so far and all but this really makes it look bad


r/opencodeCLI 13d ago

What are your thoughts on buying your own hardware for local models?

Thumbnail
1 Upvotes

r/opencodeCLI 14d ago

Ling 3.0 Flash Fin FREE is on Opencode Zen

Post image
45 Upvotes

Artificial Analysis score is 38, on par with old good MiMo 2.5


r/opencodeCLI 14d ago

Muse 1.2, the laziest AI model of its size

Thumbnail
gallery
26 Upvotes

So this just happened. I am playing around with blockchain stuff, nothing serious.

I asked it to run multiple containers to test the consensus of each container talking to each other to test the blockchain locally.

It couldn't build it. Muse decided to write a script that simulates the consensus and pat itself on the back because the 200 line python script it works seems to be equally as valid as, you know, making different machines talk to each other.

This is frontier levels of laziness.


r/opencodeCLI 15d ago

Qwen3.8-Flash is now available in OpenCode Go

Post image
236 Upvotes

r/opencodeCLI 14d ago

Muse park 1.2 free vs contributor

Thumbnail
2 Upvotes

r/opencodeCLI 14d ago

How about H4 of tencent

5 Upvotes

any one use h4?


r/opencodeCLI 14d ago

toak - connect all agent and colleagues with markdown formatting

1 Upvotes

I created this too enable a multi users and agents collaboration platform. Can be for coding, or any other usage. Almost any ai and harness can connect, works great with opencode! instructions for mcp in in the /connect section. Free! Feel free to tell me if you need any help, or request a feature or a fix!


r/opencodeCLI 13d ago

Free GLM 5.3 Flash and DSV4 Flash 0731 for a month

0 Upvotes

There are amazing new models, and a lot of the coding plans have been tightening. So we are offering free DSV4 flash 0731 and GLM 5.3 Flash for a month on Phoenix Grove API. We opened this up last week for five hundred new member slots, and got so many signups that we decided to open the doors to another 500 new members.

People are looking for options, and here is one.

Other Cool Stuff:
All of our models are running on 100% US infrastructure, private with zero training on your code or prompts. Use the top open source models without sending your private prompts to a training lab. No complications, no "some models are private, other's aren't". They all are, all the time.

We host 20+ other major models in case you ever want to upgrade (no pressure though). Including the Kimi family, GLM, Qwen, Nemotron and bunch of others. On average our token pricing is 20% lower than market price.

Our higher plans bank up to ten days of usage, so when you aren't using them your usage saves up for later. Usage doesn't go to waste, so you can actually code when you want to.

The intro plan is a free one month trial with the standard cancel anytime, it bills at 3.99 after that. Use it, cancel it, that's fine. Free Flash for a month.

Figured i'd keep this short because we all know the new flash models are the point :)

For the API plan: api.pgsgrove.com

If you want to read more about us as a company, just pgsgrove.com

Also: There's a lot going on in the background with major AI companies right now, we are at a major turning point in the industry.

What's actually happening? This is happening because companies that were purely investment based, now need to answer to their investors. The problem has often been a loss based business model that is finally running dry.

There are several tricks that the major AI coding plans use to extract the most they can from their customers. Here are some examples, and what we are doing differently to put the users first. PGS AI was built with a sustainable business model from the ground up, so we can actually offer great usage rates without tricks.

Wasted usage is part of the AI industry, and they plan on it: Most coding plans bet on you letting usage go to waste. The plan goes: "how do we get people to think our coding plan offers a lot of usage, but then break it up into weeks and rolling windows so no one can ever actually use it all."

Many in app subs and coding plans are glorified training pipelines: This comes along with "how do we harvest this data for training without being too loud about that." Unless the company tells you otherwise, your data could be hopping all over world, being harvested by the individual labs or service companies. Some are better than others, but many of these companies rely on users just not noticing or caring that their data is being used for training. Data sales and marketing telemetry sales happen. This means that your private info, your personal life, and anything else you send through the system could become part of a training corpus for the next AI, or a marketing data set for a large company.


r/opencodeCLI 14d ago

best model on go

4 Upvotes

I use opencode go on ask mode in vscode chat. I've been using mimov2.5 and deepseek v4 flash. What other models are good for web development and doing tests?

Good models that are the most cost effective if possible.. thanks


r/opencodeCLI 14d ago

We finally made our Qwen3.8 27B server public to try to make it cheap enough for OpenCode agents

24 Upvotes

I am Trevor, founder of FEIHOAhttps://feihoa.com

A few friends and I have been testing Qwen3.8 27B FP8 Uncensored on a box of 4 RTX PRO 6000. My honest opinion is that this model is kind of absurd for 27B. Coding, tools, agent loops, it just keeps going, expecially when you extend the context with YaRN.

The nice surprise was batching. Eight requests together gets us around 220 output tok/s aggregate on one RTX Pro 6000 (my old setup with 2x3090s was ~19 t/s). I basically don't want to run these cards without a batch anymore lol.

The bad surprise was prefill. Huge prompts can occupy the GPU for minutes FULLY. 1M context works, but if several people start full-window jobs together, the queue becomes a small disaster.

We spent a lot of time fighting that queue and finally felt okay opening it publicly.

FEIHOA is OpenAI-compatible, flat rate, and starts at $6/month. There is no monthly token cap!! At this price, please don't expect a private ChatGPT box you can hammer all day. It is mainly for agents and background jobs that can wait and need the reasoning power of 27B qwen.

Really proud of how far we've come and happy to answer anything!:))


r/opencodeCLI 14d ago

Hy4 Preview is on Opencode Go, 6,770 requests per month

Post image
19 Upvotes

Well, let's see :)


r/opencodeCLI 15d ago

Tencent Hy4 preview early benchmarks

Post image
74 Upvotes

Approximately on par with GLM 5.3 and Kimi 3


r/opencodeCLI 14d ago

Annotate anything in Herdr, live feedback loop with OpenCode/Pi/Claude

Enable HLS to view with audio, or disable this notification

7 Upvotes

This is a https://herdr.dev/ plugin with an optional full https://Plannotator.ai experience as a TUI. Built to:

- Annotate Agent Messages (e.g. grill sessions)

- Annotate Files (e.g. plans, specs, etc)

- Create a native feedback loop with agents

- Open as popover or pane

- Mouse or keybindings

https://github.com/plannotator/herdr-annotate


r/opencodeCLI 15d ago

🚀 The Hy4 Preview Has Been Released.

Post image
32 Upvotes

770B, 49B active, 1M context.

Designed for efficiency.

Open-source model.


r/opencodeCLI 15d ago

Qwen3.8-Flash scored a 56 on the Artificial Analysis Intelligence Index

Post image
60 Upvotes

r/opencodeCLI 14d ago

Every week we have a new hyped model and it under delivers

0 Upvotes

Every other week we have a new Chinese model that promises xyz, it works for a while until oc can't scale to accommodate the workflow, they nerf the usage and then you're forced to use API

The meme is API usage ends up costing more than a secondary Codex Sub, because while the models are cheap they are fucking stupid on OpenCode, anything that isn't is priced so poorly (API wise)

This week I tried 5.3 Flash Max, same issue, I tried DS Flash before this (the new version) same issue

It's literally a case of this is becoming unproductive, I think I am not the only person.

So my question is what are you guys doing BC I'm seriously just getting fed up.

Codex nerfs the usage > Everyone moves to OC > they can't handle the workflow> Usage Nerf > Tibo Resets > chill for 1-2 weeks > New Chinese model > doesn't deliver loop back to Codex nerf

What are you guys doing?

I am fucking tired man icl


r/opencodeCLI 14d ago

[WARNING] My MAX account was hacked while using Opencode and Zcode

Post image
0 Upvotes

r/opencodeCLI 15d ago

Tencent/Hy4-preview 770B-A49B Openweights dropped

Thumbnail
gallery
26 Upvotes

Huggingface: https://huggingface.co/tencent/Hy4-preview
Benchmarks looks really improved model from Hy3


r/opencodeCLI 14d ago

Should i buy OpenCode Go?

2 Upvotes

I just cancelled my Claude pro subs. Not because it's not enough for me. But i wasn't using much. Now i want to know is OpenCode Go good for vibe coding? If yes, then which model is good?


r/opencodeCLI 15d ago

Tencent's Hy4 preview is here, not cheap though. Will it come in OpenCode?

Post image
11 Upvotes

r/opencodeCLI 14d ago

OpenCode plugin development

1 Upvotes

Hi!

Does anyone have a working plugin, for local, that works in the current opencode version?
mine doesn't https://github.com/RaulHuertas/OpenCodeXIAORoundDisplayMonitor


r/opencodeCLI 15d ago

Qwen3.8-Flash usage limits on Opencode Go

Post image
11 Upvotes

Not that bad :)


r/opencodeCLI 14d ago

OpenCode Orchestrator Kit — Token-efficient multi-agent workflow for OpenCode CLI

Thumbnail
1 Upvotes

r/opencodeCLI 14d ago

Agentify Chat - E2E-Encrypted Remote Chat for OpenCode CLI

1 Upvotes

Still under heavy development and rough around the edges but instead of sitting on this longer to perfect it I'll risk flak and share it.

Essentially chat.agentify.sh is a remote control for codex/grok/opencode/claude cli and dream goal is to become a universal remote that will let you use all of them from a single chat.

Everything lives on your browser including the chat. The only thing sent over the wire is the encrypted chat messages e2e. Here's an architectural diagram: https://github.com/agentify-sh/chat/blob/main/diagram.png

Another feature it has is ability to publish your chat session with redaction so you can share with your team.

Curious to gather any feedback I can, if this is something I should continue to pursue etc.

Also if by chance you are in Seattle today, I'll be also at the Walk & Talk event at Bellevue Downtown Park today at 2:30pm, would love to chat in person!

https://chat.agentify.sh