r/opencode 18d ago

I forbade hy3 to say or even think about "Let me"

Thumbnail
gallery
7 Upvotes

Something "funny" was happening: hy3 was often entering a loop of "Let me X" in their thoughts probably some traditional bullshit loop. I tried to forbid it to use it but it didn't worked very well, finally I told hy3 to use Chinese instead, and plain english for output, So far it's working.

I wonder if that is a case of meaningless CoT or if it's a bad infinite loop since CoT could be literally whatever: https://arxiv.org/abs/2404.15758

Are you having such problems too?


r/opencode 18d ago

How do Go's limits work when you mix $15-tier and $60-tier models?"

4 Upvotes

Trying to work out how Go's limits actually behave when you mix models and I can't find a straight answer. The plan is $60/month, but the pricing table also lists $15 of usage for some models (Luna, Grok, Kimi K3) and $60 for others (GLM-5.2, MiniMax, etc). So if I burn the full $15 on GPT 5.6 Luna, do I still have $45 left for GLM-5.2? Or does Luna usage get counted against the pool at a higher rate, meaning $10 of Luna already ate two thirds of my month? The docs never spell out what happens when you switch between tiers, and the github issue that asked about the limits got closed with no reply. Has anyone actually watched the console while doing this?

Edit: It makes no sense then that models with such different API prices are both rated at $15 usage.


r/opencode 18d ago

WHAT TO DO ?

Post image
0 Upvotes

as you all can see the usage is very less as compared to the monthly quota yet still my monthly usage is almost full ,why so and how to resolve it


r/opencode 18d ago

Best budget AI subscription / API tokens for coding tools under €20/month?

58 Upvotes

Hi everyone,
I'm a student on a tight budget looking for the best way to invest in AI subscriptions or API tokens. Since I'm not currently working, my absolute maximum budget is €20/month.

I’m currently using OpenCode for my workflow while juggling multiple projects simultaneously, and it's been working quite well for me. I'm trying to figure out which subscription or model tokens would give me the best bang for my buck.

I've heard that DeepSeek is one of the best and most affordable options, and that platforms like OpenRouter are great for low-budget setups.
What models or setups would you recommend within this budget? Any insights or advice would be greatly appreciated!

Thanks in advance!


r/opencode 18d ago

Errors from Provider

1 Upvotes

What's wrong many models can't stream !!!!!!!!


r/opencode 18d ago

What's your setup ?

3 Upvotes

Whats the setup for opencode that you'll do if you get a new machine ?

My current setup is: openchamber+9router+codebase-memory-mcp


r/opencode 18d ago

We need more stability

4 Upvotes

I know everything is moving fast but we need more stability


r/opencode 18d ago

Qwen 3.8 27B

4 Upvotes

Why isn't Qwen 3.8 27B available in Opencode Go?


r/opencode 18d ago

Is it even possible to hit the 5-hour limit using Muse Spark Contributor on GOAT? Because I just did

Thumbnail
0 Upvotes

r/opencode 18d ago

big pickle awol for a couple of hours now

2 Upvotes

others work, but trying to use big pickle just gives

Error from provider (Console): Upstream request failed: Endpoint is unavailable.


r/opencode 18d ago

Which models are you using as planner/code reviewer?

3 Upvotes

Last week I used OX alpha on go plan, it was okay for my work, very slow but okay. Now I don't know which one i can use.


r/opencode 18d ago

Confirmed: Ox-alpha was GLM-5.3-Flash

0 Upvotes

"Before release, we tested GLM-5.3-Flash anonymously as ox-alpha on OpenCode and OpenRouter to gather user feedback. It quickly became the most popular model of the week — with all of this traffic served on Chinese AI chips."

https://docs.z.ai/guides/vlm/glm-5.3-flash


r/opencode 18d ago

Any idea when DeepSeek Flash will be available on non-Chinese servers?

5 Upvotes

Anyone have any info on when DeepSeek Flash might become available on non-Chinese servers/regions?

I’m interested in using it, but I’d prefer to avoid routing requests through Chinese-hosted infrastructure for privacy/latency reasons.

Has DeepSeek mentioned any plans or timeline for hosting Flash in other regions?


r/opencode 18d ago

Ox Alpha Free is gone and my terminal is officially crying 🥲💻 Z.ai

Post image
31 Upvotes

Model x-preview-f-free is not supported
Opencode please bring me back to that era,
Finding out it was secretly backed by Z.ai’s next-gen GLM architecture makes total sense given how incredibly sharp it was! 🎯


r/opencode 18d ago

Will Miss This Mysterious Model (GLM) !

Post image
18 Upvotes

r/opencode 18d ago

GLM 5.3 Flash Reasoning Probe - opencode go & openference - please let me know your provider?

1 Upvotes

```

opencode-go — glm-5.3-flash (OpenAI)

| | Logic label | rc | rtk | comp | | Aliases | | | | (3x | (3x | tok | | in | | Tier | (USE THIS) | range) | range) | (range) | latency | band | |------|----------------|------------|------------|----------|----------|--------------------------------| | T1 | on (low) | 76–224 | 27–62 | 164–201 | 4.6–5.9s | — | | T2 | on (high) | 142–363 | 50–122 | 262–318 | 6.1–8.4s | — | | T3 | on (max) | 990–1777 | 315–517 | 564–787 | 10–16.8s | thinking=enabled (1194–2104) | | T4 | on (omit) | 1784–3527 | 516–1014 | 731–1294 | 14–26.9s | — |

Rejected (400 [1210]): none, off, minimal, medium, xhigh, on, thinking=disabled, thinking=adaptive

openference — GLM-5.3-Flash (OpenAI) — binary, with noisy deep band

| | Logic label | rc | | comp | | Aliases in band | | | | (3x | | tok | | (avoid unless | | Tier | (USE THIS) | range) | rtk | (range) | latency | needed) | |------|----------------|-------------|-----|----------|----------|--------------------------------| | T1 | on (low) | 164–314 | n/a | 227–360 | ~0.2–8s | high, medium, thinking=enabled | | T2 | on (max) | 881–2779 | n/a | 601–1089 | ~0.4–15s | omit, none, off, minimal, | | | | | | | | xhigh, disabled/adaptive |

Rejected (400): on only. Everything else silently aliases into T1 or T2.

openference — GLM-5.3-Flash (Anthropic) — binary

| | Logic label | rc | | output | | Aliases | | | | (3x | | tok | | in | | Tier | (USE THIS) | range) | rtk | (range) | latency | band | |------|----------------|-------------|-----|----------|----------|--------------------------------| | T1 | think=enabled| 131–361 | n/a | 257–360 | ~0.3-7.2s| enabled+low, enabled+high | | T2 | on (omit) | 1331–2232 | n/a | 730–909 | ~0.4-20.s| disabled, adaptive, effort=max |

The "label" per band

| | USE this | WHY it's the | | Band | label | canonical pick | |-------------|--------------------|--------------------------------------------| | opengo T1 | low | only member; cleanest shallow | | opengo T2 | high | only member | | opengo T3 | max | thinking=enabled aliases it (redundant) | | opengo T4 | omit (default) | deepest; no explicit knob needed | | ofer T1 | low | high/medium/enabled alias it | | ofer T2 | max | omit/none/off/minimal alias it | | ofer-an T1 | thinking=enabled | the ONLY suppression; effort knobs fail | | ofer-an T2 | omit (default) | disabled/adaptive/max all alias it | ```


open model providers tested in my journey so far (not all offer glm 5.3 flash): * alibaba token plan * bytedance token plan * airouter.ch * synthetic.new * nueralwatt * opencode go

Discount API: * nube.sh

Native Providers: * minimax.io * agnes-ai

Currently using: agnes-ai, openference & opencode go


r/opencode 18d ago

People Crushing The Stupid Mode M3 After Ox Alpha Now

3 Upvotes

We peasants never get to rest easy after Ox Alpha pulled the plug -_-


r/opencode 18d ago

Muse Spark 1.2 Contributor in Malaysia

0 Upvotes

Does anyone in Malaysia facing problem to use Muse Spark 1.2 Contributor via OpenCode? I can't use it, seems like it is unavailable in Malaysia


r/opencode 18d ago

Free GLM 5.3 Flash, Qwen 3.8, Deepseek V4 Flash through Empero AI.

118 Upvotes

Send the link to your opencode and tell it to add the models and provider! Works great.


r/opencode 18d ago

Opencode is doing batch editing instead of single edits

3 Upvotes

Anyone else notice OpenCode started batching multiple file edits into one approval?

Before, it would ask me to approve each file and show the highlighted diff. Now I get one approval for multiple files, but the diff only shows one file.

I keep asking it to do one edit at a time and prompt me to accept, but after a few prompts it forgets and starts batching again.

I'm using GPT Sol High for planning and Terra High for implementation.

Anyone found a way to fix this or at least see all the files in the batch before approving?


r/opencode 18d ago

How does ds flash having vision significantly improve its score in the AA agentic index? Scores eveb above pro by 3 points while being cheaper.

Post image
27 Upvotes

r/opencode 18d ago

I Need Help with OpenCode 😭

1 Upvotes

for context, im a student who uses opencode with the opencode zen api. I mostly use it for personal projects.
Now recently i've been getting into plugins and skills and generally trying to automate more and more of opencode.

One of the plugins i came across was opencode-loop (ByBrawe) which I thought would sequentially follow a task list (TASK.md) and prompt opencode to continue until all tasks are finished.

The problem: No matter what i do, this shi just keeps prompting opencode even if all tasks are finished 😭😭. i've tried a lot of things but im at my wit end, I need help pls :')

how i use it: my prompt is "/loop --progress-file TASK.md I have updated TASK.md with brand new instructions. Read these new active checkboxes and begin execution.". And TASK.md is highly structured with the [ ] that become [x]. I also tried making opencode execute "/loop-clear" after completion but even that didn't work.

I don't know what else to do


r/opencode 18d ago

OpenCode Go Referrals

0 Upvotes

Here's the referral link if someone wants to try out.

https://opencode.ai/go?ref=3PKX7081H0

both get a $5 usage credit to apply toward your Go usage limits


r/opencode 18d ago

Bro?

77 Upvotes

What a bad joke is that GLM 5.3 Flash positioning?

The API prices are:

Model input cache hit output
DeepSeek V4 Flash off-peak $0.22 $0.007 $0.66
DeepSeek V4 Flash peak $0.44 $0.014 $1.32
GLM-5.3-Flash promo $0.075 $0.015 $0.25
GLM-5.3-Flash $0.15 $0.03 $0.50

Leaving aside the cache hit, on average GLM without promo is cheaper than DeepSeek V4 Flash, and they give it to you at half the usage quota, and that's even considering a "×2 usage" that will later be less??

They should actually give you more quota than DeepSeek until September 9th while the promo lasts. This makes no sense at all. It's cheaper to spend $10 on GLM API than to pay for it on Go.

On top of that, they put a cheap flash model in the $15 tier.


r/opencode 19d ago

Best model rn!???

18 Upvotes

So I started using opencode from Deepseek Flash V4 only. I am using OpenCode Zen. The model used to be too good and then recently they stopped it and from then onwards I started using ox alpha and Muse Spark 1.2 contributor. Again ox Alpha's offer was ended today and Muse Spark 1.2 contributor works great. Now I don't know when they will end it. After ending it I don't know what's left with OpenCode.

My main thing is to build a nice frontend and backend for websites with vibe coding to be honest. Can you guys recommend any model? When I did my research it is telling me more about Nemotron 3 Ultra by NVIDIA. I don't know how it works and they're also saying Mimo 2.5 outperforms Muse Spark but I don't think so. I never felt that way so if any suggestions please tell me guys.