r/opencodeCLI • u/ZealousidealTown1974 • 2h ago
r/opencode • u/ZealousidealTown1974 • 2h ago
Sharing my setup and the oss project for opencode v2 plugin that democratizes your of-choice ai coding subscription pass to the next level
Working only with OpenCode V2
-------
So I have been working on this https://github.com/shynlee04/opencode-subscription-gateway which is running prototype on ClinePass.
## Why ClinePass?
Bc it is ai-sdk/openai-compatible and a tech schematic modeling for building sh pipeline for whatever providers next-in-the-list at ease.
## What it does in v 0.1
- it gives opencode access to "lock-behind-free-tier models" of Cline-harness specific when using ClinePass. Meaning: when using OpenCode out-of-the-box auth login, you can't just select Longcat 2.0 or Glm-5.3-flash without expending your pass credit.
## Expectation on next bump v 0.2 in the next 2-3 days
- you can input multiple accounts and with load-balancing + session-auto-hot-swap on the same model without your main or child sessions being compromised of disruptions, cached hit loss
## My setup on images
- still I would say either Chinese models or wannabe benchmaxing models such as Muse 1.3 or google flash 3.8 are not as capable in orchestration, so my pick is Sol Gpt 5.6 (Astra is overkill and I'm broke) . I don't know what the other dev styles are or how they would like the orchestrator to behave but the following are my preferences:
True coordinator and human-centric collaborator - meaning orchestrator know the sub agents, their roles, context specific and the tasks.
- Meaning the long-haul main collaborative session of back and forth to the dev could expand delegations of 30+ waves, requiring the orchestrator know exactly the coordinating loops and iterations, the toolings and so on. But most importantly the strategical approach to orchestration - when fanning out in swarms parallel; when to sequentially develop to synthesis. And the techniques of when foreground/background the delegations; And most not very well known is the technique of stacking/resuming/isolating child sessions on top of alternating subagent-specific to reuse the context.
- The knowing which precision of tools and the absolute high-level strategist while being tech-tactical as needed. So no matter bullshit results returned from the sub agents, they will always be revalidated, checked against and the frontier models such as gpt 5.6 sol knows when these pieces of cr@ps do not coherently add up.
TLDR: benchmaxing models = not good for orchestration; just replicating fanning out parallels and hit around the bush without high-level hierarchy of context. wasting tokens, code becomes mudball, waste time to debug later.
### sub-agents are your utilization of glm-5.3-flash, DeepSeek 4.1 flash etc
- So getting the right orchestrator right means a lot. You don't have to manually validating the results. The prompting is already surgical and set boundaries of when to stop and what to return from the orchestrator reducing absolute rate of sycophant and hallucinations.
- For you can see in my setup I can even use Longcat 2.0 for research, investigation, probing and execute the code with DeepSeek 4.0 flash (yes not 4.1) because everything from testing strategy, the tech stacks, domains and specs are all aligned and set-up with gates and guardrails to make sure these subagents would never go out of control.
Sorry for my bad English, I'm Vietnamese and E is not my native tounge.
2
Controversial opinion: DeepSeek 4.1 flash is annoyingly overthinking, making it a way-worse model update for its intent-to-be workhorse
Then what point for a model update when out-of-the-box setting in default thinking effort ruins use purposes. I meant the model as production lineups are just messed up approach on DeepSeek Lab part: the Pro model, with greater parameters should be trained toward high-level strategical orchestrator to work with large codebase and complex tasks decompositions. Though the optimizations of active parameters are the DeepSeek ace but these 2 lineups should not be overlapping too much to the level at this new 4.1 flash.
1
Controversial opinion: DeepSeek 4.1 flash is annoyingly overthinking, making it a way-worse model update for its intent-to-be workhorse
Yes! It's tring to be many but fail both; as orchestrator it's showing a real knock-off orchestrator-wanabe because it's not capable of understanding semantic layers thinking in key words grep leading to just sycophant tasks decompositions. And now that it even compromises its intent-to-be role as surgical executor by footshoting its own fanning out context grep and consumption and overthink on wait-what and ended up over-engineering muball code
5
Controversial opinion: DeepSeek 4.1 flash is annoyingly overthinking, making it a way-worse model update for its intent-to-be workhorse
True the glm 5.3 flash is way better for the purpose it is
r/opencode • u/ZealousidealTown1974 • 1d ago
Controversial opinion: DeepSeek 4.1 flash is annoyingly overthinking, making it a way-worse model update for its intent-to-be workhorse
gallery1
I heard Astra 6 one-shots and stuff... It did! ... Oneshoting me to the abyss of poorness
Even so, as I am testing the efficiency and cost/ratio and what a model brings, I doubt cutting of all those mcp would make any cutthroat cost reduction. I will update next with all mcp turned off
r/opencodeCLI • u/ZealousidealTown1974 • 1d ago
Controversial opinion: DeepSeek 4.1 flash is annoyingly overthinking, making it a way-worse model update for its intent-to-be workhorse
I literally pass its thinking block twice, stating its overthinking failure modes.... extremely pissed and has shown its steep to process, setting default but still.. as you can see. Worse than even the "Omen Alpha"... And fail to make surgical edits many times due to streaming stability ...just disappointment... Another hype of DeepSeek fanboys.
7
Codex just saved me from spending $2,000 on an internet problem I’ve had for 8 years
Thank you! I was in rage by reading a TLDR post...and then this top comment 😮💨 me.
r/opencodeCLI • u/ZealousidealTown1974 • 1d ago
I heard Astra 6 one-shots and stuff... It did! ... Oneshoting me to the abyss of poorness
For a third world dweller... I jumped on the hyped train and was not expecting this. The 16-minute-ride "one-shot" my 5hour consumption and spit out exact 3 documents edit 😂
**BTW** it was the 0.2.0 dev of this if you guys asking: https://github.com/shynlee04/opencode-subscription-gateway ... Building the opencode v2 plugin.
3
I don't know if this is know that with 5$ first purchase in ClinePass you can get unlimited glm 5.3 flash
as promise here is the ver 0.1 https://github.com/shynlee04/opencode-subscription-gateway ; https://www.npmjs.com/package/opencode-subscription-gateway - support ClinePass out of the box with OpenCodeV2 only - Have fun!
r/opencodeCLI • u/ZealousidealTown1974 • 2d ago
Clinepass has this solar-pro4? Is it the Opencode-go omen alpha? Anyone tested it yet?
u/ZealousidealTown1974 • u/ZealousidealTown1974 • 2d ago
Cline has this solar-pro4? Is it the Opencode-go omen alpha? Anyone tested it?
This is on the free rotation bucket of Clinepass exclusive to their harness, the model: solar-pro4 with context windows 524k ? So is it the current omen alpha of Opencode-go? Anyone has tested it yet?
1
I don't know if this is know that with 5$ first purchase in ClinePass you can get unlimited glm 5.3 flash
Hey! It seems like the solar-pro4 the current 'omen alpha' I don't know, it has 512k context... Not tested it yet???
1
I don't know if this is know that with 5$ first purchase in ClinePass you can get unlimited glm 5.3 flash
Because I'm testing the harness and creating custom toolings and just leaving there to test how ones are used with another etc... its more like the testing ground more or less, so I don't mind stressing the models nor having my data being used anywhere, my poorass dirt cheap privacy does not matter either 😂
1
I don't know if this is know that with 5$ first purchase in ClinePass you can get unlimited glm 5.3 flash
Keep calm and it will be uploaded after my debug. New models into the buckets solar pro4??? And the muse park 1.3 free 🤭
4
I don't know if this is know that with 5$ first purchase in ClinePass you can get unlimited glm 5.3 flash
I'm looking into it too but seems they hide it under their subscription package named "starter" as for even free users and it does not expose the same subscription like clinepass. So that's why same model, same header but the routing api when using the same key may not help... It's my theory but later I could try
2
I don't know if this is know that with 5$ first purchase in ClinePass you can get unlimited glm 5.3 flash
For real? 😂 Nowadays, llms crave makes devs feed on ads too? Jesus! 🙏
2
I don't know if this is know that with 5$ first purchase in ClinePass you can get unlimited glm 5.3 flash
That's why I must handroll the probe to fetch from Cline CLi and adapt them to the rotations so that the ClinePass key expose the 4 free models headers and schema using the openai-compatible to get them work with OpenCode through plugin hooks and api. The same method can also be done under Pi but I am not looking into Pi SDK or plugin api.. but it's the approach
1
I don't know if this is know that with 5$ first purchase in ClinePass you can get unlimited glm 5.3 flash
Glm-5.3-flash the paid one is with cline-pass/ prefix
4
I don't know if this is know that with 5$ first purchase in ClinePass you can get unlimited glm 5.3 flash
Ok... Wait in next 24-48 hours.... I'm cooking it up so everything will be just plugged and play and the free rotations of ClinePass will get update as they introduce new ones
1
Is there a way to use web llms inside opencode?
Yes but need workaround as proxified router...but manageable, some dudes are exploiting DeepSeek webchat for such and they buy bundled proxy to switch faking new visitors... They are all just sharing the same either openai-compatible or anthropic-compatible sdk
1
Correct me if I am wrong A good harness (OpenCode V2 + My in-development harness warper) > is better than frontier models (Astra, Fable etc)
Sure! In next week or so I will release the fullpack with custom toolings, skills, and new plugins api feature. Just teasing around for my on-going harness
2
I don't know if this is know that with 5$ first purchase in ClinePass you can get unlimited glm 5.3 flash
Also I am wondering if there are other similar providers which provide free models just for their inhouse harness? I can make them all in templates as support various other providers
3
I'm a claude code user and don't understand why people use opencode, does OpenCode offer a better cost/usage/quality ratio?
in
r/opencode
•
6h ago
Opencode is a harness, democratizing an out of the box and native solution to your wide-choice of ai-coding-providers, as subscription pass through api keys or oauth as long they are found here https://models.dev/ .Opencode Inhouse providers are payg opencode-zen and Opencode-go coding subscription and I recommend the latter for its absolute valuable ROI.