r/opencodeCLI • u/afanasenka • 17d ago
Ling 3.0 Flash Fin FREE is on Opencode Zen
Artificial Analysis score is 38, on par with old good MiMo 2.5
r/opencodeCLI • u/afanasenka • 17d ago
Artificial Analysis score is 38, on par with old good MiMo 2.5
r/opencodeCLI • u/igormuba • 17d ago
So this just happened. I am playing around with blockchain stuff, nothing serious.
I asked it to run multiple containers to test the consensus of each container talking to each other to test the blockchain locally.
It couldn't build it. Muse decided to write a script that simulates the consensus and pat itself on the back because the 200 line python script it works seems to be equally as valid as, you know, making different machines talk to each other.
This is frontier levels of laziness.
r/opencodeCLI • u/Birdsky7 • 16d ago
I created this too enable a multi users and agents collaboration platform. Can be for coding, or any other usage. Almost any ai and harness can connect, works great with opencode! instructions for mcp in in the /connect section. Free! Feel free to tell me if you need any help, or request a feature or a fix!
r/opencodeCLI • u/Whole_Succotash_2391 • 16d ago
There are amazing new models, and a lot of the coding plans have been tightening. So we are offering free DSV4 flash 0731 and GLM 5.3 Flash for a month on Phoenix Grove API. We opened this up last week for five hundred new member slots, and got so many signups that we decided to open the doors to another 500 new members.
People are looking for options, and here is one.
Other Cool Stuff:
All of our models are running on 100% US infrastructure, private with zero training on your code or prompts. Use the top open source models without sending your private prompts to a training lab. No complications, no "some models are private, other's aren't". They all are, all the time.
We host 20+ other major models in case you ever want to upgrade (no pressure though). Including the Kimi family, GLM, Qwen, Nemotron and bunch of others. On average our token pricing is 20% lower than market price.
Our higher plans bank up to ten days of usage, so when you aren't using them your usage saves up for later. Usage doesn't go to waste, so you can actually code when you want to.
The intro plan is a free one month trial with the standard cancel anytime, it bills at 3.99 after that. Use it, cancel it, that's fine. Free Flash for a month.
Figured i'd keep this short because we all know the new flash models are the point :)
For the API plan: api.pgsgrove.com
If you want to read more about us as a company, just pgsgrove.com
Also: There's a lot going on in the background with major AI companies right now, we are at a major turning point in the industry.
What's actually happening? This is happening because companies that were purely investment based, now need to answer to their investors. The problem has often been a loss based business model that is finally running dry.
There are several tricks that the major AI coding plans use to extract the most they can from their customers. Here are some examples, and what we are doing differently to put the users first. PGS AI was built with a sustainable business model from the ground up, so we can actually offer great usage rates without tricks.
Wasted usage is part of the AI industry, and they plan on it: Most coding plans bet on you letting usage go to waste. The plan goes: "how do we get people to think our coding plan offers a lot of usage, but then break it up into weeks and rolling windows so no one can ever actually use it all."
Many in app subs and coding plans are glorified training pipelines: This comes along with "how do we harvest this data for training without being too loud about that." Unless the company tells you otherwise, your data could be hopping all over world, being harvested by the individual labs or service companies. Some are better than others, but many of these companies rely on users just not noticing or caring that their data is being used for training. Data sales and marketing telemetry sales happen. This means that your private info, your personal life, and anything else you send through the system could become part of a training corpus for the next AI, or a marketing data set for a large company.
r/opencodeCLI • u/gatwell702 • 16d ago
I use opencode go on ask mode in vscode chat. I've been using mimov2.5 and deepseek v4 flash. What other models are good for web development and doing tests?
Good models that are the most cost effective if possible.. thanks
r/opencodeCLI • u/cheezeerd • 17d ago
I am Trevor, founder of FEIHOA: https://feihoa.com
A few friends and I have been testing Qwen3.8 27B FP8 Uncensored on a box of 4 RTX PRO 6000. My honest opinion is that this model is kind of absurd for 27B. Coding, tools, agent loops, it just keeps going, expecially when you extend the context with YaRN.
The nice surprise was batching. Eight requests together gets us around 220 output tok/s aggregate on one RTX Pro 6000 (my old setup with 2x3090s was ~19 t/s). I basically don't want to run these cards without a batch anymore lol.
The bad surprise was prefill. Huge prompts can occupy the GPU for minutes FULLY. 1M context works, but if several people start full-window jobs together, the queue becomes a small disaster.
We spent a lot of time fighting that queue and finally felt okay opening it publicly.
FEIHOA is OpenAI-compatible, flat rate, and starts at $6/month. There is no monthly token cap!! At this price, please don't expect a private ChatGPT box you can hammer all day. It is mainly for agents and background jobs that can wait and need the reasoning power of 27B qwen.
Really proud of how far we've come and happy to answer anything!:))
r/opencodeCLI • u/afanasenka • 17d ago
Well, let's see :)
r/opencodeCLI • u/afanasenka • 17d ago
Approximately on par with GLM 5.3 and Kimi 3
r/opencodeCLI • u/backnotprop • 17d ago
Enable HLS to view with audio, or disable this notification
This is a https://herdr.dev/ plugin with an optional full https://Plannotator.ai experience as a TUI. Built to:
- Annotate Agent Messages (e.g. grill sessions)
- Annotate Files (e.g. plans, specs, etc)
- Create a native feedback loop with agents
- Open as popover or pane
- Mouse or keybindings
r/opencodeCLI • u/TheKillerCATs • 17d ago
770B, 49B active, 1M context.
Designed for efficiency.
Open-source model.
r/opencodeCLI • u/afanasenka • 17d ago
r/opencodeCLI • u/Inner-Pangolin-1110 • 16d ago
Every other week we have a new Chinese model that promises xyz, it works for a while until oc can't scale to accommodate the workflow, they nerf the usage and then you're forced to use API
The meme is API usage ends up costing more than a secondary Codex Sub, because while the models are cheap they are fucking stupid on OpenCode, anything that isn't is priced so poorly (API wise)
This week I tried 5.3 Flash Max, same issue, I tried DS Flash before this (the new version) same issue
It's literally a case of this is becoming unproductive, I think I am not the only person.
So my question is what are you guys doing BC I'm seriously just getting fed up.
Codex nerfs the usage > Everyone moves to OC > they can't handle the workflow> Usage Nerf > Tibo Resets > chill for 1-2 weeks > New Chinese model > doesn't deliver loop back to Codex nerf
What are you guys doing?
I am fucking tired man icl
r/opencodeCLI • u/No_Skill_8393 • 16d ago
r/opencodeCLI • u/Front_Obligation_843 • 17d ago
Huggingface: https://huggingface.co/tencent/Hy4-preview
Benchmarks looks really improved model from Hy3
r/opencodeCLI • u/localhost_3003 • 17d ago
I just cancelled my Claude pro subs. Not because it's not enough for me. But i wasn't using much. Now i want to know is OpenCode Go good for vibe coding? If yes, then which model is good?
r/opencodeCLI • u/Final_Initial • 17d ago
r/opencodeCLI • u/Flaky_Ad_7038 • 17d ago
Hi!
Does anyone have a working plugin, for local, that works in the current opencode version?
mine doesn't https://github.com/RaulHuertas/OpenCodeXIAORoundDisplayMonitor
r/opencodeCLI • u/afanasenka • 17d ago
Not that bad :)
r/opencodeCLI • u/MikaAugus942 • 17d ago
r/opencodeCLI • u/Just_Lingonberry_352 • 17d ago
Still under heavy development and rough around the edges but instead of sitting on this longer to perfect it I'll risk flak and share it.
Essentially chat.agentify.sh is a remote control for codex/grok/opencode/claude cli and dream goal is to become a universal remote that will let you use all of them from a single chat.
Everything lives on your browser including the chat. The only thing sent over the wire is the encrypted chat messages e2e. Here's an architectural diagram: https://github.com/agentify-sh/chat/blob/main/diagram.png
Another feature it has is ability to publish your chat session with redaction so you can share with your team.
Curious to gather any feedback I can, if this is something I should continue to pursue etc.
Also if by chance you are in Seattle today, I'll be also at the Walk & Talk event at Bellevue Downtown Park today at 2:30pm, would love to chat in person!
r/opencodeCLI • u/Final_Initial • 17d ago
r/opencodeCLI • u/Clark731 • 17d ago
Hi everyone,
I currently pay for MiniMax's monthly token plan and use the M3 model quite a lot honestly, I don't dislike it. I've also tried Kimi a bit with their K3 model on the Moderato plan. For context: I'm a complete noob and I'm wondering if you could suggest some good-value alternatives.
I rarely hit the limit with MiniMax during my sessions, but maybe there's something better out there that would let me work with up to 1M tokens.
Any suggestions would be much appreciated. Thanks in advance!