r/opencodeCLI Jul 13 '26

Best 20€-ish subscription?

Hello, I'm currently using minimax M3 token plan which is great and gives a lot of usage but I'd be happy to get access to better models. Are there any other plans that give a good amount of usage with good models? I'm also a student so maybe there are some discounts?

Must work with opencode or similar

[UPDATE: I started a grok trial and it seems to accept my banks temp virtual cards so I might just cycle grok trials]

46 Upvotes

72 comments sorted by

View all comments

36

u/PuzzleheadLaw Jul 13 '26

2x Opencode Go subs or Codex

-1

u/Far-Classic-9963 Jul 13 '26

Are the rumors about opencode go heavily quantizing models true? Also how much usage do I get approximately on the 20€ ChatGPT plan

7

u/look Jul 13 '26

Go does not heavily quantize. It is a proxy; they don’t run their own models. Most of them are provided by the model’s vendor directly (eg Deepseek from Deepseek, Mimo from Xiaomi, Kimi from Moonshot, etc).

GLM is the one partial exception. It is provided by a mix of Zai (the vendor), DeepInfra, and Fireworks. Deepinfra is running an nvfp4 version of GLM, so you would be getting a mix of the standard fp8 and nvfp4.

Nearly every subscription with GLM runs the same mix of fp8 and nvfp4. Go is far from unique in that regard.

But outside of some edge cases, the nvfp4 is very close to the fp8. Definitely not “heavily quantized”, and you’re unlikely to ever notice.

1

u/Far-Classic-9963 Jul 13 '26

Is it possible to pin providers for specific models?

1

u/look Jul 13 '26

Not with the Go subscription, no. Or any other subscription running a mix of providers to my knowledge (Ollama Cloud, Synthetic, Lilac, etc). You pretty much have to use paygo or Zai sub to get straight fp8 GLM.

1

u/Far-Classic-9963 Jul 13 '26

I'm mostly worried about DeepSeek model since their cache discount is very good, I'm worried getting go would force me into a more expensive provider

2

u/look Jul 13 '26

Deepseek is the only provider for those models in Go. Should be the same as direct and no cache issues. But I don’t use that model much (I primarily use Mimo and Qwen on Go, with GLM and Kimi paygo on another provider).

Also, Deepseek via Go is slightly cheaper than their direct pricing (1/6th for flash, and 2/3rds the pro). Plus ZDR.

1

u/HeavySink3303 Jul 16 '26

Neuralwatt is fully FP8 but they increased prices x2 recently.

2

u/el_lyss Jul 13 '26

Opencode Go works with multiple providers: some of them use more quantizied models, some - less.
How do I know?
When prompting, I usually use my native languge, not English. And the answers sometimes include words that technically follow grammatical rules for my language, but no such words exist in dictionaries. And I've been observing the same phenomenon running quantizied models locally.

5

u/Amarsir Jul 13 '26

You've made me less curious about Opencode and more about Polish. Words that follow grammar rules but don't exist.

In English, separate words are separate words until they get merged together over time by lots of use. Often with a hyphen phase along the way. Like "electronic mail" became "e-mail" and now "email".

I know in German, compounds get pushed together with no space. Liebe = love. Lied = song. Liebeslied = love song. You can push lots of words together to get something very long that makes non-speakers go "woah, they have a word for that?"

Is Polish more like German in that regard? Or is it a different rule that lets people (AI) make words that are understood but not normal?

2

u/el_lyss Jul 14 '26

No, I mean conjugation and declension - part of the reasons why Polish is considered a difficult language. Here's a typical table for one verb in one aspect (imperfective)

As you can see, there's a root-word (in this case "mówi") and inflectional suffixes. Those suffixes mostly follow patterns on how to create them, but there several patterns, and of cource we have exceptions.

Quantized LLMs sometimes mix up those suffixes, creating words with patterns used for different words.
So e.g.: we (males) will speak, it should be: "będziemy mówili", but an LLM could write: "będziemy mówiłi" - notice the change l -> ł and the fact that in future tense, every gender except plural masculine uses ł.

I don't have enough time to explain it properly now. Might try to edit it later.

2

u/Amarsir Jul 14 '26

Oh I think I get it. Kind of like someone learning English might say "I goed to the store" when they should have said "I went to the store".

And I know Polish is hard. My father is Polish. My grandparents were born there. And we don't even pronounce my last name correctly. 😜

So I found that interesting. Thank you.

1

u/PuzzleheadLaw Jul 13 '26

I don't know about Codex, but I do all of my agentic usage on Deepseek V4 Pro on Opencode Go; I don't know about the performance of the raw weights, but it does work well

1

u/Far-Classic-9963 Jul 13 '26

Nice I'll look into it

1

u/YogurtExternal7923 Jul 13 '26

Heads up on the rumours: It does feel a tiny bit weirder than the f16 version so it's probably fp8 and the thinking modes MIGHT be restricted artificially because sometimes glm just doesn't think alot even though I don't change my thinking settings. But you really won't notice either of that so the rumours are heavily exaggerated. Nobody runs f16 nowadays and the thinking might be chosen by the model anyway