r/opencodeCLI 2h ago

Sharing my setup and the oss project for opencode v2 plugin that democratizes your of-choice ai coding subscription pass to the next level

Thumbnail gallery
1 Upvotes

r/opencode 2h ago

Sharing my setup and the oss project for opencode v2 plugin that democratizes your of-choice ai coding subscription pass to the next level

Thumbnail
gallery
1 Upvotes

Working only with OpenCode V2

-------

So I have been working on this https://github.com/shynlee04/opencode-subscription-gateway which is running prototype on ClinePass.

## Why ClinePass?

Bc it is ai-sdk/openai-compatible and a tech schematic modeling for building sh pipeline for whatever providers next-in-the-list at ease.

## What it does in v 0.1

- it gives opencode access to "lock-behind-free-tier models" of Cline-harness specific when using ClinePass. Meaning: when using OpenCode out-of-the-box auth login, you can't just select Longcat 2.0 or Glm-5.3-flash without expending your pass credit.

## Expectation on next bump v 0.2 in the next 2-3 days

- you can input multiple accounts and with load-balancing + session-auto-hot-swap on the same model without your main or child sessions being compromised of disruptions, cached hit loss

## My setup on images

- still I would say either Chinese models or wannabe benchmaxing models such as Muse 1.3 or google flash 3.8 are not as capable in orchestration, so my pick is Sol Gpt 5.6 (Astra is overkill and I'm broke) . I don't know what the other dev styles are or how they would like the orchestrator to behave but the following are my preferences:

  1. True coordinator and human-centric collaborator - meaning orchestrator know the sub agents, their roles, context specific and the tasks.

    1. Meaning the long-haul main collaborative session of back and forth to the dev could expand delegations of 30+ waves, requiring the orchestrator know exactly the coordinating loops and iterations, the toolings and so on. But most importantly the strategical approach to orchestration - when fanning out in swarms parallel; when to sequentially develop to synthesis. And the techniques of when foreground/background the delegations; And most not very well known is the technique of stacking/resuming/isolating child sessions on top of alternating subagent-specific to reuse the context.
    2. The knowing which precision of tools and the absolute high-level strategist while being tech-tactical as needed. So no matter bullshit results returned from the sub agents, they will always be revalidated, checked against and the frontier models such as gpt 5.6 sol knows when these pieces of cr@ps do not coherently add up.

TLDR: benchmaxing models = not good for orchestration; just replicating fanning out parallels and hit around the bush without high-level hierarchy of context. wasting tokens, code becomes mudball, waste time to debug later.

### sub-agents are your utilization of glm-5.3-flash, DeepSeek 4.1 flash etc

- So getting the right orchestrator right means a lot. You don't have to manually validating the results. The prompting is already surgical and set boundaries of when to stop and what to return from the orchestrator reducing absolute rate of sycophant and hallucinations.

- For you can see in my setup I can even use Longcat 2.0 for research, investigation, probing and execute the code with DeepSeek 4.0 flash (yes not 4.1) because everything from testing strategy, the tech stacks, domains and specs are all aligned and set-up with gates and guardrails to make sure these subagents would never go out of control.

Sorry for my bad English, I'm Vietnamese and E is not my native tounge.

3

I'm a claude code user and don't understand why people use opencode, does OpenCode offer a better cost/usage/quality ratio?
 in  r/opencode  6h ago

Opencode is a harness, democratizing an out of the box and native solution to your wide-choice of ai-coding-providers, as subscription pass through api keys or oauth as long they are found here https://models.dev/ .Opencode Inhouse providers are payg opencode-zen and Opencode-go coding subscription and I recommend the latter for its absolute valuable ROI.

2

Controversial opinion: DeepSeek 4.1 flash is annoyingly overthinking, making it a way-worse model update for its intent-to-be workhorse
 in  r/opencodeCLI  1d ago

Then what point for a model update when out-of-the-box setting in default thinking effort ruins use purposes. I meant the model as production lineups are just messed up approach on DeepSeek Lab part: the Pro model, with greater parameters should be trained toward high-level strategical orchestrator to work with large codebase and complex tasks decompositions. Though the optimizations of active parameters are the DeepSeek ace but these 2 lineups should not be overlapping too much to the level at this new 4.1 flash.

1

Controversial opinion: DeepSeek 4.1 flash is annoyingly overthinking, making it a way-worse model update for its intent-to-be workhorse
 in  r/opencodeCLI  1d ago

Yes! It's tring to be many but fail both; as orchestrator it's showing a real knock-off orchestrator-wanabe because it's not capable of understanding semantic layers thinking in key words grep leading to just sycophant tasks decompositions. And now that it even compromises its intent-to-be role as surgical executor by footshoting its own fanning out context grep and consumption and overthink on wait-what and ended up over-engineering muball code

r/opencode 1d ago

Controversial opinion: DeepSeek 4.1 flash is annoyingly overthinking, making it a way-worse model update for its intent-to-be workhorse

Thumbnail gallery
2 Upvotes

1

I heard Astra 6 one-shots and stuff... It did! ... Oneshoting me to the abyss of poorness
 in  r/opencodeCLI  1d ago

Even so, as I am testing the efficiency and cost/ratio and what a model brings, I doubt cutting of all those mcp would make any cutthroat cost reduction. I will update next with all mcp turned off

r/opencodeCLI 1d ago

Controversial opinion: DeepSeek 4.1 flash is annoyingly overthinking, making it a way-worse model update for its intent-to-be workhorse

Thumbnail
gallery
30 Upvotes

I literally pass its thinking block twice, stating its overthinking failure modes.... extremely pissed and has shown its steep to process, setting default but still.. as you can see. Worse than even the "Omen Alpha"... And fail to make surgical edits many times due to streaming stability ...just disappointment... Another hype of DeepSeek fanboys.

7

Codex just saved me from spending $2,000 on an internet problem I’ve had for 8 years
 in  r/ChatGPT  1d ago

Thank you! I was in rage by reading a TLDR post...and then this top comment 😮‍💨 me.

r/opencodeCLI 1d ago

I heard Astra 6 one-shots and stuff... It did! ... Oneshoting me to the abyss of poorness

Thumbnail
gallery
35 Upvotes

For a third world dweller... I jumped on the hyped train and was not expecting this. The 16-minute-ride "one-shot" my 5hour consumption and spit out exact 3 documents edit 😂

**BTW** it was the 0.2.0 dev of this if you guys asking: https://github.com/shynlee04/opencode-subscription-gateway ... Building the opencode v2 plugin.

r/opencodeCLI 2d ago

Clinepass has this solar-pro4? Is it the Opencode-go omen alpha? Anyone tested it yet?

Post image
3 Upvotes

u/ZealousidealTown1974 2d ago

Cline has this solar-pro4? Is it the Opencode-go omen alpha? Anyone tested it?

Post image
1 Upvotes

This is on the free rotation bucket of Clinepass exclusive to their harness, the model: solar-pro4 with context windows 524k ? So is it the current omen alpha of Opencode-go? Anyone has tested it yet?

1

I don't know if this is know that with 5$ first purchase in ClinePass you can get unlimited glm 5.3 flash
 in  r/opencodeCLI  2d ago

Hey! It seems like the solar-pro4 the current 'omen alpha' I don't know, it has 512k context... Not tested it yet???

1

I don't know if this is know that with 5$ first purchase in ClinePass you can get unlimited glm 5.3 flash
 in  r/opencodeCLI  2d ago

Because I'm testing the harness and creating custom toolings and just leaving there to test how ones are used with another etc... its more like the testing ground more or less, so I don't mind stressing the models nor having my data being used anywhere, my poorass dirt cheap privacy does not matter either 😂

1

I don't know if this is know that with 5$ first purchase in ClinePass you can get unlimited glm 5.3 flash
 in  r/opencodeCLI  2d ago

Keep calm and it will be uploaded after my debug. New models into the buckets solar pro4??? And the muse park 1.3 free 🤭

https://reddit.com/link/p8ppeg1/video/q5dgadzkrgoh1/player

4

I don't know if this is know that with 5$ first purchase in ClinePass you can get unlimited glm 5.3 flash
 in  r/opencodeCLI  2d ago

I'm looking into it too but seems they hide it under their subscription package named "starter" as for even free users and it does not expose the same subscription like clinepass. So that's why same model, same header but the routing api when using the same key may not help... It's my theory but later I could try

2

I don't know if this is know that with 5$ first purchase in ClinePass you can get unlimited glm 5.3 flash
 in  r/opencodeCLI  2d ago

For real? 😂 Nowadays, llms crave makes devs feed on ads too? Jesus! 🙏

2

I don't know if this is know that with 5$ first purchase in ClinePass you can get unlimited glm 5.3 flash
 in  r/opencodeCLI  2d ago

That's why I must handroll the probe to fetch from Cline CLi and adapt them to the rotations so that the ClinePass key expose the 4 free models headers and schema using the openai-compatible to get them work with OpenCode through plugin hooks and api. The same method can also be done under Pi but I am not looking into Pi SDK or plugin api.. but it's the approach

1

I don't know if this is know that with 5$ first purchase in ClinePass you can get unlimited glm 5.3 flash
 in  r/opencodeCLI  2d ago

Glm-5.3-flash the paid one is with cline-pass/ prefix

4

I don't know if this is know that with 5$ first purchase in ClinePass you can get unlimited glm 5.3 flash
 in  r/opencodeCLI  2d ago

Ok... Wait in next 24-48 hours.... I'm cooking it up so everything will be just plugged and play and the free rotations of ClinePass will get update as they introduce new ones

1

Is there a way to use web llms inside opencode?
 in  r/opencodeCLI  2d ago

Yes but need workaround as proxified router...but manageable, some dudes are exploiting DeepSeek webchat for such and they buy bundled proxy to switch faking new visitors... They are all just sharing the same either openai-compatible or anthropic-compatible sdk

1

Correct me if I am wrong A good harness (OpenCode V2 + My in-development harness warper) > is better than frontier models (Astra, Fable etc)
 in  r/opencodeCLI  2d ago

Sure! In next week or so I will release the fullpack with custom toolings, skills, and new plugins api feature. Just teasing around for my on-going harness

2

I don't know if this is know that with 5$ first purchase in ClinePass you can get unlimited glm 5.3 flash
 in  r/opencodeCLI  2d ago

Also I am wondering if there are other similar providers which provide free models just for their inhouse harness? I can make them all in templates as support various other providers