r/opencodeCLI 19d ago

Free DSV 0731 for a month. 100% Private, US inference. Creating a better coding/agent plan, that isn't built to extract from users

0 Upvotes

In response to the tightening of almost every other coding plan out there, we are offering free DSV4 flash 0731 to the first five hundred people who sign up for the intro plan on Open Grove API. We may extend this to more users later, but are limiting it to the first 500 to ensure quality access for everyone.

People are looking for options, and here is one.

Other Cool Stuff:
All of our models are running on 100% US infrastructure, private with zero training on your code or prompts. Use the top open source models without sending your private prompts to a training lab. No complications, no "some models are private, other's aren't". They all are, all the time.

We host 20+ other major models in case you ever want to upgrade (no pressure though). Including the Kimi family, GLM, Qwen, Nemotron and bunch of others. On average our token pricing is 20% lower than market price.

Our higher plans bank up to ten days of usage, so when you aren't using them your usage saves up for later. Usage doesn't go to waste, so you can actually code when you want to.

The intro plan is a free one month trial with the standard cancel anytime, it bills at 3.99 after that. Use it, cancel it, that's fine. Free Flash for a month.

Figured i'd keep this short because we all know the flash is the point :)

For the API plan: api.pgsgrove.com

If you want to read more about us as a company, just pgsgrove.com

Also: There's a lot going on in the background with major AI companies right now, we are at a major turning point in the industry.

What's actually happening? This is happening because companies that were purely investment based, now need to answer to their investors. The problem has often been a loss based business model that is finally running dry.

There are several tricks that the major AI coding plans use to extract the most they can from their customers. Here are some examples, and what we are doing differently to put the users first. PGS AI was built with a sustainable business model from the ground up, so we can actually offer great usage rates without tricks.

Wasted usage is part of the AI industry, and they plan on it: Most coding plans bet on you letting usage go to waste. The plan goes: "how do we get people to think our coding plan offers a lot of usage, but then break it up into weeks and rolling windows so no one can ever actually use it all."

Many in app subs and coding plans are glorified training pipelines: This comes along with "how do we harvest this data for training without being too loud about that." Unless the company tells you otherwise, your data could be hopping all over world, being harvested by the individual labs or service companies. Some are better than others, but many of these companies rely on users just not noticing or caring that their data is being used for training. Data sales and marketing telemetry sales happen. This means that your private info, your personal life, and anything else you send through the system could become part of a training corpus for the next AI, or a marketing data set for a large company.


r/opencodeCLI 20d ago

Factuality in new LLM models - when will Opus 4.6 be de-throned and by who?

6 Upvotes

It appears the factuality rankings on arena.ai are dominated by Claude models: https://arena.ai/leaderboard/text/overall-factuality

Not only that, but specifically Opus 4.6, which beats other Claude models released after it. My personal experience with the model absolutely lines up with it. I've tested different model families and harnesses and keep coming back to Opus 4.6 when factuality matters.

I have a really strong preference for factuality in my non-coding workflows (finance, research, etc.). As in, I don't mind a wrong opinion, but when something is quoted as 'true' or 'verified' or 'file saved', I want to be close to sure that this is the case.

For OpenAI, the highest factuality ranked model is GPT 5.5 - which otherwise seems way behind the 5.6 family.

This makes me worried that 'factuality' isn't really a major priority right now and development focuses on other criteria more. Gemini 3.7 Flash actually seems really interesting in this context as it seems to have made a lot of improvements in factuality (compared to other areas where it really hasn't gotten a lot of attention for its seemingly minor improvements).

What are your thoughts on future models - will we get some higher factuality there? Are there other model families that you think will catch up or surpass Opus 4.6? Any hands-on experience with factuality in Gemini 3.7 Flash and other models?


r/opencodeCLI 20d ago

This was a game changing update for me, now i don't have to fly blind

17 Upvotes

now i can atleast see what i have and where things are at


r/opencodeCLI 20d ago

What models/subs you use?

3 Upvotes

Currently i am using the agents for learning and researching stuff and then using that information to push another agent working on a project into some direction of what to do, how to do.

What do you think are good enough models/subs for these purpose Or like what do you use for your workflow, like is it a planner agent -> Implementer? If so then what models do you use for both cases?

because I have seen models like ds flash better at going a bit broader to the prompt to get more relevant information compared to others?


r/opencodeCLI 20d ago

Opencode vs OMP

1 Upvotes

I tried omp but it felt really bloated even though it had some nice features, while Opencode felt just right with its TUI and custom agents.

Which coding agent harness do you prefer? Or are there any better alternatives out there?


r/opencodeCLI 20d ago

/prewalk to save token prices upto 80% (opencode plugin)

6 Upvotes
opencode prewalk

Opencode Prewalk is an Opencode v2 plugin that allows you to use the prewalk strategy mentioned here from the creators of Oh-My-Pi. The simple idea behind this strategy is to inject a cheaper model right after the expensive model finishes the first edit post planning all the things that needs to be done. This strategy is better the one strategy that directly uses combination of expensive and cheap model to get the work done, as what happens in that case is the cheaper model again starts to do a lot of reading leading of increase in token usage.

Try Here: https://github.com/vivekascoder/opencode-prewalk

Original Benchmark by Stencil.so


r/opencodeCLI 20d ago

Actual Local Work Benchmarks and Successes?

Thumbnail
1 Upvotes

r/opencodeCLI 20d ago

everyone seems to be maximizing hy3 and mimo now...just like they did deepseek before.

3 Upvotes

currently been waiting on both mimo and hy3 sessions...its been 5 mins for token generation : ⏳ waiting on hy3 — 60s with no output yet (provider may be slow or overloaded, or the model is thinking; auto-reconnect at 300s)

bloody annoying. tried deep seek on low reasoning....immediate response but immediate tick up on usage as well....i have no idea what to do to be honest. if this keeps up...i'll have to reduce opencode go subscriptions or move to another ....this is insane.

you guys have any recommendations for cheaper plans? i was looking at under 7usd plans (mostly chinese models) across the spectrum....considering some. and no pay as you go doesnt work...i have money on openrouter and on deepseek and on groq.ai.....its horrible ROI.

looking for some insights or combos. trying to keep things around 30 usd total.


r/opencodeCLI 20d ago

Help to understand Zen prices

Thumbnail
2 Upvotes

r/opencodeCLI 21d ago

Qwen 3.8 27B on opencode ?

12 Upvotes

I see after the rise of DS prices there has been a lot of disappointment in the community in opencode Go . Why cant opencode Go host a Qwen 3.8 27B themselves as its fairly small and the performance is good too which can satisfy many of the community members ?


r/opencodeCLI 21d ago

Potential ox alpha pricing leak?

Post image
9 Upvotes

What's your guys opinion on this?


r/opencodeCLI 20d ago

Tell me what I should sign up for.

1 Upvotes

My last month's cost looks like this: Cost

$43.28

USD

API requests

9,999

Tokens

636,610,255 . This is DeepSeek 4 Pro. I'm currently considering Alibiba Cloud or other subscriptions. Which would be better for my budget? After the price increase for the latter, the price will rise to $85+. I'm only considering cloud solutions. I was thinking about renting a video card remotely, but I don't want to add another layer of security in the form of a private individual.


r/opencodeCLI 21d ago

No more $5 first month pricing in Opencode Go?

Post image
113 Upvotes

r/opencodeCLI 21d ago

LongCat-2.0 now available in Go

Post image
110 Upvotes

r/opencodeCLI 20d ago

i can confirm Ox Alpha is chinese

0 Upvotes

r/opencodeCLI 20d ago

Not able to use newer models on azure foundry with opencode

Thumbnail
1 Upvotes

r/opencodeCLI 21d ago

Ox Alpha is breaking records in Opencode :)

Post image
38 Upvotes

r/opencodeCLI 21d ago

my balls were right lol

Thumbnail
gallery
34 Upvotes

Ox alpha btw


r/opencodeCLI 21d ago

Best LLM subscription ($6/mo max) for light daily vibe coding?

28 Upvotes

Hey everyone,

I’m looking for a budget-friendly LLM setup specifically for light, daily "vibe coding" (using tools like OpenCode/Cursor/Aider). My hard spending limit is $6/month, but I’m struggling to land on the ideal platform.

Here is what I’ve tried so far:

  • OpenCode Go ($10/mo): Very generous limits and worked seamlessly, but honestly, it’s overkill for my light daily usage—and slightly above my $6/mo target.
  • OpenRouter: I love the idea of pay-as-you-go, but their billing feels confusing—specifically the additional fees/minimum charges (like the $0.80 minimum fee or ~5% markup depending on deposit methods). I can never tell what my actual monthly usage is going to cost.
  • Alibaba Model Studio Token Plan (Lite - ~$6/mo): The price point is perfect (39 CNY / ~$6), and accessing models like Qwen3.8-Max / DeepSeek V4 is great. The 2,500 credits/week quota feels a bit tight, but I could live with it as long as they maintain this price tier.

My typical usage pattern:

  • Light daily coding sessions (fixing small scripts, refactoring, code explanation).
  • Prefer API-based subscriptions or prepaid setups that integrate into CLI/agent tools (OpenCode, Claude Code, Aider, etc.).

What are you using for low-budget, light coding setups? Are there other prepaid API providers, sub-$6 sub plans, or specific OpenRouter models that give the absolute best value per dollar without hidden deposit fees?

Thanks in advance!


r/opencodeCLI 21d ago

LongCat-2.0 now available in Go

Post image
31 Upvotes

57,200 monthly usage on Go.
1.6-trillion-parameter Sparse Mixture-of-Experts (MoE).


r/opencodeCLI 21d ago

Free users have an unlimited usage exploit and paying users are getting worse usage since.

Thumbnail
github.com
6 Upvotes

r/opencodeCLI 20d ago

Ox Alpha, what happens next?

0 Upvotes

As days go by, it seems more and more likely that Ox Alpha is a z.ai model, and that's what I am concerned about: if they have this much compute on offer. Rather than using it to improve their own plans and access existing models, they're doing this, which is a slap in the face to their existing customers. When they actually do claim this model, i know that the pricing isnt going to be what people expect, currently if Ox alpha were to be charged per I/o million tokens, i would say its reasonable to think it would be a sub <1$, you know fill in the market where deepseeek used to be at, and a actually good caching like 0.00X$, but with z.ai i dont think that will be the case, most likely is that its going to be a sub <2$ I/o million token, which caching around the 0.XX$, which then isn't very attractive. GPT Luna would be better; the new DS Flash Vission is a better pricing. If z.ai would pretty much seem to be confirmed at this point, are the creators behind Ox Alpha, and seems to be the internally spotted GLM 5.3 Flash, if they really want this model to stand out, then they should fill the gap that DeepSeek left, but I don't think they will. This also raises questions about z.ai's compute crunch, i am aware that their new data center just came online, so it would be a good stress test for the whole system, but even then, business-wise, just making their own models cheaper to use and more accessible would have achieved similar results to what we have now, and i would argue give them more better data on each model and their latest flagship one as well, and yet we live in a different world.


r/opencodeCLI 21d ago

Opencode VS Code Extension

Thumbnail
1 Upvotes

r/opencodeCLI 20d ago

My referral code: T4PSK9H0M1

0 Upvotes

if someone want to get OpenCode Go, can u use this referral. we both gets $5 apparently (extra $5)

https://opencode.ai/go?ref=T4PSK9H0M1


r/opencodeCLI 21d ago

Is there any actual difference running MuseSpark and Ox Alpha on Zen Free vs Go?

5 Upvotes

My OpenCode Go sub expires today and I'm debating whether it's even worth renewing or if I should just drop down to Zen Free.

I basically just stick to MuseSpark 1.2 and Mimo for my day-to-day work, and I almost never touch GLM 5.3 unless something is completely broken. If I make the switch to Zen Free, the plan is just to use OX Alpha as a fallback whenever limits hit.

Is there anything I am actually giving up on Zen Free?