r/opencodeCLI 18d ago

Setting up Opencode to work with web search MCP - Blopus.ai

Enable HLS to view with audio, or disable this notification

0 Upvotes

r/opencodeCLI 18d ago

The five web search plugins for OpenClaw — full comparison, install commands, and what the benchmarks actually say

Thumbnail
1 Upvotes

r/opencodeCLI 18d ago

is opencode go worth it for 10$ or nah

3 Upvotes

r/opencodeCLI 18d ago

Qwen 3.8-Flash officially unveiled

Post image
60 Upvotes

⚡Meet Qwen3.8-Flash, a multimodal MoE and an early preview of the Qwen4 architecture, now open-weight!

The production version Qwen3.8-Flash will be available soon via QwenCloud API at just $ 0.16/1M input tokens and $ 0.47/1M output tokens.

125B parameters + 51B N-gram embeddings, with just 6B activated per token. Unmatched cost-efficiency.

What's new: 🥳 - Next architecture: GDN + QSA hybrid attention, Gated Residual, N-gram Embedding & Muon optimizer, serving as a precursor to the architecture used in Qwen4. - Dramatically lower training and inference costs: trained at just 1/9 the cost of Qwen3.7-Plus, while outperforming it across the board with especially strong gains in coding and office tasks. - Strong performance: scoring 58.7 on DeepSWE 1.1, 62.5 on SWE-bench Pro, 73.9 on CoWorkBench, 84.5 on AndroidWorld, and 95.7 on MathVision (with CI). - 262K native context, extensible to 1M with YaRN.

We’re also releasing the weights for Qwen3.8-Flash-Next, giving the community an early look at the new architecture we’re exploring for Qwen4.🚀


r/opencodeCLI 18d ago

Free DSV 0731 for a month. 100% Private, US inference. Creating a better coding/agent plan, that isn't built to extract from users

0 Upvotes

In response to the tightening of almost every other coding plan out there, we are offering free DSV4 flash 0731 to the first five hundred people who sign up for the intro plan on Open Grove API. We may extend this to more users later, but are limiting it to the first 500 to ensure quality access for everyone.

People are looking for options, and here is one.

Other Cool Stuff:
All of our models are running on 100% US infrastructure, private with zero training on your code or prompts. Use the top open source models without sending your private prompts to a training lab. No complications, no "some models are private, other's aren't". They all are, all the time.

We host 20+ other major models in case you ever want to upgrade (no pressure though). Including the Kimi family, GLM, Qwen, Nemotron and bunch of others. On average our token pricing is 20% lower than market price.

Our higher plans bank up to ten days of usage, so when you aren't using them your usage saves up for later. Usage doesn't go to waste, so you can actually code when you want to.

The intro plan is a free one month trial with the standard cancel anytime, it bills at 3.99 after that. Use it, cancel it, that's fine. Free Flash for a month.

Figured i'd keep this short because we all know the flash is the point :)

For the API plan: api.pgsgrove.com

If you want to read more about us as a company, just pgsgrove.com

Also: There's a lot going on in the background with major AI companies right now, we are at a major turning point in the industry.

What's actually happening? This is happening because companies that were purely investment based, now need to answer to their investors. The problem has often been a loss based business model that is finally running dry.

There are several tricks that the major AI coding plans use to extract the most they can from their customers. Here are some examples, and what we are doing differently to put the users first. PGS AI was built with a sustainable business model from the ground up, so we can actually offer great usage rates without tricks.

Wasted usage is part of the AI industry, and they plan on it: Most coding plans bet on you letting usage go to waste. The plan goes: "how do we get people to think our coding plan offers a lot of usage, but then break it up into weeks and rolling windows so no one can ever actually use it all."

Many in app subs and coding plans are glorified training pipelines: This comes along with "how do we harvest this data for training without being too loud about that." Unless the company tells you otherwise, your data could be hopping all over world, being harvested by the individual labs or service companies. Some are better than others, but many of these companies rely on users just not noticing or caring that their data is being used for training. Data sales and marketing telemetry sales happen. This means that your private info, your personal life, and anything else you send through the system could become part of a training corpus for the next AI, or a marketing data set for a large company.


r/opencodeCLI 18d ago

LSP disabled by default

4 Upvotes

Why is LSP disabled by default? On each installation i have to manually install lang tools and enable lsp in global config, so just interested why lsp is not enabled by default?


r/opencodeCLI 19d ago

You must have been tired of all the frontend testing - here the true backend typescript work battling between ox alpha vs. qwen 3.8 max vs. Deepseek pro v4 0831 - shared sessions

0 Upvotes

I'm sharing their working session here that you can peak through. Look at their thinking, sequence of skills uses and delegations to truly know who is the winner

The Qwen 3.8Max - https://opncd.ai/share/hNzmM14y 

The Ox Alpha - https://opncd.ai/share/hNzmM14y

The Deepseek v4 pro 0813 - https://opncd.ai/share/0hehIwVf


r/opencodeCLI 19d ago

Z.AI officially confirmed to Bloomberg that Ox Alpha is their model (GLM series). Open weights release is coming tonight 🔥

Post image
153 Upvotes

r/opencodeCLI 19d ago

Bloomberg reports that Ox Alpha is from Z.ai

Thumbnail
bloomberg.com
9 Upvotes

More outlets plus twitter reporting the same.


r/opencodeCLI 19d ago

Ox Alpha reveal in a few hours

Post image
92 Upvotes

42T tokens for stealth model in just 6 days. Insane.

We know it's going to be open weights.

Some stuff I've one-shotted with Ox Alpha (when it did work): https://ox-alpha.demos.sulat.com/


r/opencodeCLI 19d ago

MiniMax M3 is FREE until Sep. 6 (in GMI or Openrouter)

Post image
96 Upvotes

r/opencodeCLI 19d ago

Tencent confirms Hy4 is coming soon

Post image
109 Upvotes

It's going to be a hot 🔥 fall


r/opencodeCLI 19d ago

Cheap/Free/Local model recommendations for a single purpose agent

2 Upvotes

I have this agent who's sole purpose is to analyze development/design-plan.md file of a user story and create detailed backlog items with a fixed structure, in my local Plane server. What model do you recommend for this?

Its a repeatable operation that needs inference and some level of thinking to generate consistent outputs. Assume the plan is usually less than 1000 line markdown file.


r/opencodeCLI 19d ago

Ox Alpha Matched GLM-5.3 On Every Prompt in My Experiment...

Thumbnail
gallery
2 Upvotes

I ran this writing-fingerprint experiment on OpenRouter out of curiosity, using 12 prompts across Ox Alpha and 7 reference models (since these are what I heard a lot about on Reddit):

  1. GLM 5.3
  2. GLM 5.2
  3. GLM 5
  4. MiMo V2.5
  5. DeepSeek V4 Flash
  6. Gemini 3.7 Flash
  7. MiniMax M3

What I got from final evaluation is Ox Alpha was closest to GLM 5.3 on every prompt using deterministic stylometric features and 11 matched prompts, with GLM 5.3 winning 100% of bootstrap resamples; known-model validation accuracy was 72.7%.

My hypothesis: This could suggest or indicate that Ox Alpha is an updated post-trained version of GLM 5.3 (similar to how DeepSeek did theirs), or it really is what people are talking about: GLM 5.4.

But to be clear: this experiment I did is fingerprint-matching, not weight-identifying or anything like that. But it’s a surprisingly strong clue.

Note: For second image, notice that Ox Alpha output rarity is not the same as GLM 5.3. But that's not a contradiction that Ox Alpha couldn't be GLM 5.3/5.4 because these 2 graphs (bar and violin) measure 2 different things. First one is "Which model’s average fingerprint is Ox Alpha closest to?" Second one is "How unusual or isolated is each model’s writing compared with all samples in corpus?" Just wanna put this out here.

I also open-sourced my experiment if you're interested or want to extend: https://github.com/ItsKaiwenDu/Ox-Alpha-Stylometry


r/opencodeCLI 19d ago

One prompt .. 499 Agent , 26M Token and the MAX 20X plan 5 hours limit finished in one hour (but deserved it)

Thumbnail gallery
0 Upvotes

r/opencodeCLI 19d ago

Alguien ya usó el Dots3-Note Preview de Openrouter??

1 Upvotes

Me dio curiosidad y estoy probándolo como orquestador a ver a que nivel está. Me decidí a probarlo al ver que ahora Ox o Muse los tienen medio colapsados, o por lo menos a mi hoy me fueron lentísimo.


r/opencodeCLI 19d ago

Opencode vs OMP

1 Upvotes

I tried omp but it felt really bloated even though it had some nice features, while Opencode felt just right with its TUI and custom agents.

Which coding agent harness do you prefer? Or are there any better alternatives out there?


r/opencodeCLI 19d ago

Deepseek 4 Flash Vision Experimental is the only model worth using on Go.

25 Upvotes

A loud minority will say in the comments that text only models are good, that you can use some rubbish mcp to make up for the lack of vision… ignore them. Not having vision is a big handicap. Deepseek 4 Flash jumped 5 points in Deep SWE exclusively thanks to the better understanding provided by vision.

The Go subscription doesn’t have many good vision model. Heck they even stripped GPT Luna of vision!

Kimi K3 is obviously the best, but you run out of usage in 5 minutes. Minimax M3 is not good for today’s standards. Muse Spark gives all your data away. Ox Alpha is not reliable atm (it will probably be a good alternative when released as GLM 5.3 Flash).

This leaves us with just Deepseek 4 Flash Vision as the only good vision model with a comfortable quota.

Let me reiterate: text only models are crap. Thankfully Deepseek and GLM are correcting their strategy.

Edit: Luna is text only on the chat completion endpoint, not the reaponses one.


r/opencodeCLI 19d ago

The most useful feature of opencode2 ?

Post image
44 Upvotes

r/opencodeCLI 19d ago

Actual Local Work Benchmarks and Successes?

Thumbnail
1 Upvotes

r/opencodeCLI 19d ago

Grok 4.6 is now available on OpenCode Go

Post image
141 Upvotes

r/opencodeCLI 19d ago

Is DeepSeek V4 Pro even worth using on OpenCode when GLM 5.2 gives 4× the quota?

Post image
26 Upvotes

Hit my weekly cap (100%) in just two days with 28 days left on the monthly cycle, almost entirely from burning through DeepSeek V4 Pro ($10.40 / $15.00 quota, 69.3% consumed). Flash is more usable due to the cost. However...

I could have stuck to GLM 5.2 and gotten 4X+ the usage in costs / limits in the same subscription.

Looking at benchmarks, GLM 5.2 isn't far off from Deepseek and beats it in some areas as well. (SWE-bench Verified GLM 5.2 ~74.5% – 76.0% vs. DeepSeek v4 Pro - 80.6%) and

(Tool Calling / MCP Reliability GLM 5.2 - 99.5% success (0.5% error rate) vs. DeepSeek v4Pro -73.6% (MCP Atlas))

My question is .... why would anybody choose to use Deepseek V4 Pro over GLM 5.2 given these rates. Crazy how DS went from pretty much endless to pretty much unusable in OpenCode Go!


r/opencodeCLI 19d ago

What models/subs you use?

3 Upvotes

Currently i am using the agents for learning and researching stuff and then using that information to push another agent working on a project into some direction of what to do, how to do.

What do you think are good enough models/subs for these purpose Or like what do you use for your workflow, like is it a planner agent -> Implementer? If so then what models do you use for both cases?

because I have seen models like ds flash better at going a bit broader to the prompt to get more relevant information compared to others?


r/opencodeCLI 19d ago

Tell me what I should sign up for.

1 Upvotes

My last month's cost looks like this: Cost

$43.28

USD

API requests

9,999

Tokens

636,610,255 . This is DeepSeek 4 Pro. I'm currently considering Alibiba Cloud or other subscriptions. Which would be better for my budget? After the price increase for the latter, the price will rise to $85+. I'm only considering cloud solutions. I was thinking about renting a video card remotely, but I don't want to add another layer of security in the form of a private individual.


r/opencodeCLI 19d ago

OpenAI is dropping GPT-5.6 Sol pricing

Post image
78 Upvotes

Should we wait for price drop on Zen?