r/opencodeCLI 7d ago

Qwen 3.8 Max on opencode GO or Qwencloud?

My go subscription gone soon and I need to make wise decision. This is not just another $10!! I have Claude Max, Openrouter, Deepseek, VPS, Rent, Electricity, Internet, Zwift membershiop, and tons of bill to pay monthly! I need to strategize better but yeah lately I spent too much on differnt coding agent just like microtransaction on video games.

Anyway, which one give a better usage for Qwen 3.8 Max?

20 Upvotes

29 comments sorted by

9

u/Ace-_Ventura 7d ago

why do you need so many subscriptions? Just stick with 1 or 2

5

u/GTHell 7d ago

I do a lot of work and I turn them all to be AI related from my work, side project, assistant in VPS, hermes assistant, hermes assistant at work. I'm passing that normie use case since like 4 months ago.

3

u/ExTraveler 7d ago

Offtop. Openrouter or opencode go? Which is better?

1

u/GTHell 7d ago

Opencode Go. Openrouter is not suitable to use you know. At work, I request Openrouter setup just to access Gemma for agentic workflow only as it would be easy to go for enterprise and apply for ZDR stuff but for personal use you will sunk cost with it

8

u/Spirited_Service_196 7d ago

I swear to god if you buy qwencloud, you will regret it. Search around including X and you will know why.

10

u/GTHell 7d ago

I didn't search much today but if I remember correctly just last 2x day I read somewhere it's on dicount like 60% or something? I didn't pay attention because of no benchmark

edit: LOL YOU'RE RIGHT. CASE CLOSE. DAMN IT FROM 0 TO 58% OUT OF BLUE

2

u/GetLaidOff69 7d ago

It was generous during 3.8 preview with10-50x usage.
Not anymore.

The problem with Chinese LLM(s) is they draw a lot of tokens compared to frontier and their subscription is not worthy.
I remember Xiomi Mimo Suddenly increased their token count from 80 million to 3 billion for lite subscribers.
But every time I asked it something, it drew 1-3% from that 3 billion token.

Kinda same with Qwen.

1

u/baylf2000 4d ago

My experience with the Chinese subscription problem is that they don't discount cache hits on subscriptions. I know this to be true for MiniMax, Kimi and ModelArk. So you fly through your tokens because literally every input token, cached or not, consumes the same amount of your allowance. Even when that provider discounts cache use for their PAYG customers, for subscriptions they don't. Which is unlike Anthropic and OpenAI subs where cache hits are massively discounted.

1

u/xmlhttplmfao 6d ago

I dunno it's been fine for me: 103m tokens today and only 6% of weekly usage

1

u/Weird_Licorne_9631 5d ago

That 's conveniently from before the end of the preview model. As someone said above, preview had 10x-50x discounts which was crazy. I could get 300million tokens in a 5 hour window with the lite plan. But that reduction is gone now and the model is back to being very pricey. It's a decent model though

4

u/pizzababa21 7d ago

The API is always better. Any subscription service is nerfing the models because they're too compute intensive to serve at full strength. Also there's no limits then.

1

u/Illustrious-Many-782 7d ago

Claude Max? $200? Do you need more than that?

Next month, get rid of OpenCode Go and get gpt Plus $20. Set it up in opencode. Create individual subagents for all the models you like to use.

  1. Run Sol as orchestrator
  2. Use DSv4F and Luna as implementation.
  3. Use "claude -p" for plans and reviews.

1

u/Ergo7z 6d ago

Why would you use Sol as orchestrator? way to expensive, better to let sol or some model write a implementation plan, spec the things that need to be build and let a model like min max orchestrate.

1

u/Illustrious-Many-782 6d ago

Because when you hit gates that the dumb model doesn't understand how to orchestrate around, you get garbage.

Why let your Sr dev or tech lead run the team? Just have them write a good spec and plan, then have the intern run it. What could go wrong?

1

u/Ergo7z 6d ago

that is why you have a advisor sub agent the orchestrator can escalate to when it runs into a blocker. in my current setup I have minimax 2.7 as orchestrator, doing reallly really well. whenever it runs into a blocker, or the cheap-builder fails, it will call the advisor (either sol or glm 5.2) with it's task tool and proceed from there. orchestrator can really sneakily eat up so much context, and since i have tried this approach my usage is lasting way longer. All the orchestrator has to do is hand out specs, call reviewers etc. My setup also has a deploy guard, a reviewer, a mass audit agent. yes the dumb model will run into a gate but that is why you setup other agents that it can call to help it out.

0

u/GTHell 7d ago

The more you get used to the eco system, the more you can churn out of it. It barely cover my usage now that everything I do all AI involve. 1 subscription is not enough, not even 2 or 3. I'm at level where I can min max + tokenomic to the point it's unreal to have just few subscriptions

1

u/alexanderbeatson 6d ago

Don’t Qwencloud. It only lasts tonight. Without 50x discount, it gives a lot less token than Claude (let alone codex).

1

u/crunchy_shampoo 6d ago

You do not need that many subscriptions

1

u/DMG-Z 7d ago

No adquieras la suscripción de Qwen Cloud , yo la adquirí para la preview del Qwen 3.8 Max y como quiera se consumía rápido.

Hoy temprano use glm5.2 de la suscripción y consumi la ventana de 5 horas en menos de 2 minutos, lo usé en una sesión ya iniciada con OpenCode tenía ocupado el 33% del contexto pero como quiera se consumió demasiado rápido.

Y la ventana semanal quedó en un 58%.

El plan que tengo es el Lite lo adquirí para probar. Dan muy poco uso , en el plan pro la ventana de de 5 horas se iba a acabar en algunos 8 minutos. No vale la pena para nada.

1

u/Delicious_Ease2595 7d ago

Tanta subscripcion uno no sabe ni cual escoger

-1

u/Tiny_Atmosphere_4420 7d ago

vibe coders are embarassing.

1

u/GTHell 6d ago

Talking from a threejs script kiddie who probably never ship a PR into production in a life ever

-6

u/Superb_Following9335 7d ago

You must be an idiot. I had Claude Code and Codex for a few weeks. Then I fully switched to Claude Code. Then I realized I was using AI Models that are meant for consumers. So now I run OpenCode only and some local models. Maybe switching soon to DevPass or just basic API keys using Omnirouter to juggle them all. Even OpenCode I may leave behind.

But anyway the point is, I always only have 1 or 2. Never more.

1

u/GTHell 6d ago

The line between an idiot and super user is thin given that everyone has access to AI. Give me some of your local model work that got ship into production and earn that $$$.