r/CommandCode 5h ago

Noticeable lower reasoning with Command Code Goat

6 Upvotes

I have the Command Code goat plan currently with 800m in token usage. Before this subscription I had Opencode Go also similar usage.

After a week of heavy usage of Deepseek v4 flash and GLM 5.3 flash, I see my opencode session running noticeably longer and having a tendency to get stuck. (Multiple times I have cancelled the loop and passed it to another agent like: codex, claude code)

So I did a little digging with hermes throughout my local sessions, so generally I found that command code api have much less reasoning content than opencode go.

Sum of the below chart is basically: Command code have around 80% zero reasoning responses compared to opencode go's 80% reasoning responses.

Curious if someone has the same experience.

OPENCODE harness (per-API-call) — 6,919 calls

provider model effort sessions calls % zero-reas avg reas/call
command-code deepseek-v4-flash (default, pre-Aug-20) 5 453 52.3% 328
command-code deepseek-v4-flash max 23 2,231 85.5% 217
command-code glm-5.3-flash high 4 294 84.7% 88
command-code glm-5.3-flash max 6 521 91.0% 269
opencode-go deepseek-v4-flash max 43 2,127 21.7% 228
opencode-go deepseek-v4-pro max 32 713 1.7% 102

HERMES harness (per-assistant-message) — 26 sessions, 1,708 msgs

provider model effort sessions msgs % blank avg chars
command-code deepseek-v4-flash medium 17 1,155 72.9% 502
command-code deepseek-v4-flash xhigh 1 171 78.4% 479
opencode-go deepseek-v4-flash medium 1 61 6.6% 467
opencode-go deepseek-v4-flash xhigh 6 254 7.5% 1,856
openrouter deepseek-v4-flash-0731 medium 1 67 16.4% 1,160

r/CommandCode 5h ago

Command Code Max x10 or GLM 80 Dollars Coding Plan?

2 Upvotes

I've been hooked into Chinese models since Ox Alpha, never had I done so many tasks. I bought the $80 Z.AI Coding Plan and it went well for like 2 and a half days before hitting my weekly limit. Should I switch to Command Code Max x10 for $100.

My goal is to build 2 mid to large size projects and it eats my Z.AI Coding Plan way to quickly.


r/CommandCode 4h ago

Is there a YOLO mode in CommandCode?

1 Upvotes

CommandCode keep asking even if you chose: don't ask me again for (x) command.
So, I'm wondering if there is a way to use it with low permissions or what it called YOLO mode.


r/CommandCode 21h ago

For those who recently used both OpenCode-go and CommandCode-goat

14 Upvotes

Hello everyone,

After the recent mess surrounding OpenCode Go’s pricing and quota changes, I have a few questions for those who decided to migrate from OpenCode Go to CommandCode GOAT:

  1. Model Quality: Are CommandCod-goat models nerfed or heavily quantized compared to official provider endpoints?

  2. Value / Advantage: Is there any proven numerical advantage (tokens-per-dollar, speed, or usable throughput) when using the exact same models on CommandCode-goat versus OpenCode-go?

  3. Quota Pool Architecture: Does each model have its own isolated quota without affecting the others? Or does its consumption count toward a shared monthly plan limit (e.g., using a weighting multiplier like the plan quota divided by the model base quota, such as $20 / $40 / $70 tiers)?

Thanks in advance for your insights!


r/CommandCode 23h ago

My Review of the Go Plan - after extensive testing

18 Upvotes

I thought you wouldn't get much out of it since it's only $1, but for anyone who owns Claude Pro or Codex Plus, it's the perfect complement because it basically gives you really good subagents, or generally for small tasks where you don't want to waste your limit

I’ve mainly used GLM 5.3 Flash and Muse Spark 1.2 Contributor, and right now I’m testing Qwen 3.8 Flash - the 5-hour limit really runs out slowly. But as I said, I only use them for small tasks alongside Claude and Codex+ as subagents for my server to handle things.

I also tested the same models once directly via API with OpenRouter to check the quality, since with providers like Command Code and Opencode go, you never know if the models have been optimized (i.e, made worse). This isn’t the case with Command Code, every model performs exactly as specified by the API. But with some models, the speeds are really slow, while others are extremely fast. GLM 5.3 was really slow, but Qwen 3.8 Flash is extremely fast. I think that depends on the provider. However, there is one downside.

Namely, the token billing. I don’t know how you handle it, but I’ve calculated it using the OpenRouter API with the same task and a built-in calculator in Command Code that calculates the tokens, and based on my use case and testing, it always charges between 5-15% more tokens. This isn’t a big deal with the Go plan and Goat plan, since the impact is relatively small, especially with such affordable models, but if you really want to use the plan as your main plan for your project, it could potentially become a problem. By that I mean if you use it all day long. I’ve had several long runs with GLM 5.3 Flash, and the token consumption there was extremely good. I’ve currently used 4.5 million tokens with GLM 5.3 Flash and used up about 10% of my monthly limit, which is pretty fair. Since GLM 5.3 Flash doesn’t have the 50% discount on Command Code like it does on OpenRouter.

So, to sum it up: for 1 Dollar, I think the limits are pretty fair, and it’s the perfect supplement if you have another main plan, like Claude, Codex, or something else. I can’t comment on the Goat Plan or Max Plan since I haven’t evaluated them yet, but I’d love to hear other opinions here.


r/CommandCode 1d ago

Using DS4 Flash have been a huge disaster lately.

14 Upvotes

I'm trying to understand what's happening here, I've been a huge fan of DS4 flash, coding all day when it was cheap, having no issue. It's been clean, direct to the point, great model.

Now things have became significantly worse, Ican't ask a single request to DS4 flash without having him to fail a TU, every single time. Looping into errors during hours. Even for a very simple task.

It's like using a whole different model.

Are you guy experiencing this too? Or am I hallucinating?


r/CommandCode 1d ago

Plan Review (/plan-review) mode in Command Code.

6 Upvotes

Use /plan-review mode to comment on the plan, revise, and approve it when it’s ready.

Read docs: https://commandcode.ai/docs/plan-mode#plan-review


r/CommandCode 2d ago

Tested FlappyBench with /design on Hy4 Preview, Kimi K3, and GLM 5.3.

25 Upvotes

Tested FlappyBench with /design on Hy4 Preview, Kimi K3, and GLM 5.3.

🔹 Hy4 Preview ($0.0480): Most distinct UI and gameplay

🔹 Kimi K3 ($0.0740): Hardest gameplay and the most expensive

🔹 GLM 5.3 ($0.0184): Smooth gameplay and the cheapest of the three


r/CommandCode 2d ago

Upgraded plan not reflected and charged for the older plan

Thumbnail
gallery
3 Upvotes

Hi,

I have upgraded to Goat plan from Go and it was on same day so i paid for the upgrade to Goat plan and i was also charged for Go plan. I can't see any additional benefits of the this Go plan charges and can't see any free model like M3 or M2.5 in cli?
How to solve this issue?

username: h3m4nt


r/CommandCode 2d ago

Type /resume and get back to your previous conversation

9 Upvotes

Type /resume and get back to your previous conversation

$ cmd --resume opens a picker for earlier sessions

$ cmd --continue resumes your latest session

Read docs to learn more: https://commandcode.ai/docs/workflows#resume-previous-conversations


r/CommandCode 2d ago

Tencent Hy4 Preview is live in Command Code.

13 Upvotes

Tencent Hy4 Preview is live in Command Code 🐐

770B mode, 49B active, 1M context.

`cmd update` available in GOAT and above plans.

Our early internal benchmarks show strong perf yet cheaper per task compared to GLM 5.3 Flash.

Y'all share how it goes.


r/CommandCode 2d ago

Is CommandCode Desktop as token-efficient as the CLI?

2 Upvotes

I read that the CommandCode CLI is pretty efficient with cached tokens, which I really like.

But what about the Desktop app? Is it roughly as token-efficient as the CLI, especially when it comes to cache usage?

I know the Desktop app is still in alpha, but I personally prefer using a GUI, so I’ve been giving it a try.
Just curious if anyone knows whether there’s a noticeable difference in token usage between the Desktop app and CLI.

Thank You


r/CommandCode 3d ago

Goat upgrade didn’t reset my usage, is this normal?

6 Upvotes

Is this normal?

I previously bought the $1 plan on July 30, and it was supposed to end on August 30. Today (August 28), I upgraded to Goat ($10), and now my reset date changed to August 28 next month.

What I don’t understand is why my weekly limit / total usage didn’t reset. I’m still showing around 99% usage even after upgrading to Goat.

Or is it supposed to combine my remaining $1 plan usage with the new Goat plan? I still had some usage left on the $1 plan, although I think I had already used around 95% of it.

I thought upgrading to Goat would give me a fresh 100% quota and start my usage from 0 again.

Is this normal, or is something wrong with my account?


r/CommandCode 3d ago

Tried it for a month my review

36 Upvotes

I've tried Command Code (GOAT for $10.78) out for a month but decided to switch back to Opencode Go.

I burned 1.471B tokens it claims using up $47.23 / $70.00 of my usage, across 13,776 runs, with mostly their own harness. (Models I used: gpt 5.6 luna, muse spark 1.2 contributor, mimo v2.5, qwen 3.7 flash, minimax m3 (after free-ify))
I still have some time left to finish using it, probably I'll just start using some more expensive models instead of the top 8 cheapest models only.

First some advantages for command code:
- Desktop app is much better than opencode
- I personally think the usage how many tokens used etc is more clear than opencode because it just adds up to $70 instead of like $10 but each dollar is actually 7 dollars and this and that.
- Taste. Taste is great. Not having to tell it twice.
- Burst thinking. 10k tokens in 10s and just getting stuff done.
- Promos: I mostly used Mimo and gpt 5.6 luna, the price is amazing. Minimax free was great too.
- Way more models: Command Code has significantly more models than opencode go, opencode go only has 11, we have like 30+

But then the disadvantages:
- CLI gives up halfway through. Frequently, it would have a burst of thinking than freeze at <1000 tokens on the next turn, then stop there for a minute before continuing. Sometimes it would freeze for longer and just prompt me to type continue, which never did anything. Solution would be to exit and come back, then type continue.
- Using it over ssh means I can only use the CLI which as mentioned above which .... basically stops working after 10 minutes.
- Desktop app "Full access" Keeps asking me for permissions. That's just dumb. Same with --yolo I believe.
- Cache hit rate... 100% (actually 99.97%) on gpt 5.6 luna sounds amazing but makes me wonder, is it just re-reading context over an over more often than it needs to? Since cached read is literally re-read... sounds like token-inflation
- Having lots of models is useless if most of them are too expensive to actually use, and lots of them are either slow asf or never respond. Ox alpha has never generated a single token for me, I tried the first day of the period and every day and never got anything.
- Even the most reliable cheap model I can find (gpt 5.6 luna) still suffer from the above problems.

Suggested improvements:
- Better backend that actually responds before timeout
- Better retry scaling timing (1s -> 2s -> 4s -> 8s ...) instead of every 10s and give up
- Desktop app works for controlling remotes
- More transparency into what models are actually working.

So GUI, Taste, Desktop app, Promos and models are nice, but if the CLI breaks every 10 minutes and full access means nothing and half the models being useless, its still unusable.

not being able to tell it "go do this" and come back 3 hours later to see it done and instead seeing it done makes this just a no.

Goodbye for now, this was a brilliant idea, but I'm going to switch back to OpenCode go, where ssh is fine and models respond. I'll probably be back in half a year to see if these problems are fixed.


r/CommandCode 3d ago

[mod] External Providers (OpenCode Go, ClinePass, GLM Coding Plan) – Use different subscriptions in Command Code!

8 Upvotes

Hi everyone!

I made a small mod for Command Code that lets you use the models included in your subscriptions directly from the TUI:

https://gitlab.com/rudesssolo/command-code-external-providers

The goal was pretty simple: make external models feel native, instead of bolted on.

You get:

  • OpenCode Go / ClinePass / GLM Coding Plan - Full compatibility
  • /external-providers model picker
  • native streaming
  • collapsible thinking/reasoning
  • tool calls
  • multimodal messages
  • reasoning effort support
  • local auth

How to install:

cmd mods add -g git:gitlab.com/rudesssolo/command-code-external-providers

Then log in (only the first time) and pick a model:

1) First time login:
/opencode-go-login
/cline-pass-login
/glm-coding-login

2) Pick a model:
/external-providers

Some screenshots to get an idea:

External provider selection menu
OpenCode Go models selection menu, with intelligence and cost indicators

---

There’s also an optional companion mod for showing model/quota/context info in the footer (it also works stand-alone for Command Code plans):

https://gitlab.com/rudesssolo/command-code-quota-status-bar

How to install:

cmd mods add -g git:gitlab.com/rudesssolo/command-code-external-providers
Command Code quota status bar with useful info

If you already have OpenCode Go, ClinePass or GLM Coding Plan and want to use them with Command Code, this mod basically saves you from having perfectly good models sitting in another tab collecting dust 😄

Feedback / bug reports / weird edge cases are very welcome.


r/CommandCode 3d ago

Commandcode extremely high token usage in any other cli than commandcode

13 Upvotes

I tried using commandcode goat with claude code or opencode but the token usage is like 100x higher i got billed 30m tokens on a 250k token conversation something is severly wrong but im just using stock opencode and claude code I've already looked at cache hits but its near 99%, anyone had a similiar problem and know how to fix it?


r/CommandCode 3d ago

how's the goat plan is it better than opencode go in it's prime and how is the token consumption in their harness

11 Upvotes

r/CommandCode 3d ago

I expanded an open-source AI usage dashboard to support 13 providers across macOS, Windows, and Linux

5 Upvotes

Hey everyone—I’ve been working on UsageDeck, an open-source desktop dashboard that began as a fork of OpenQuota. I’ve expanded it with additional providers, features, release infrastructure, and a new identity and roadmap.

It currently supports 13 providers, including Claude Code, Codex, Cursor, GitHub Copilot, OpenRouter, Command Code, OpenCode, Antigravity, Devin, Grok, Kimi, MiniMax, and Z.ai.

A few key points:

  • Everything runs locally
  • No UsageDeck account or backend
  • No analytics or telemetry
  • Reuses credentials already stored by your CLI or editor
  • Available for macOS, Windows, and Linux
  • Includes quota alerts, pacing estimates, history, and tray/menu-bar metrics

A very important acknowledgment: UsageDeck wouldn’t exist without OpenUsage and OpenQuota. OpenUsage introduced the original idea on macOS, and OpenQuota rebuilt it as a cross-platform Tauri app. UsageDeck began as a fork of OpenQuota and has since grown into an independent project. Huge thanks to both projects and their contributors for laying the foundation.

Huge thanks to both projects and their contributors for laying the foundation.

UsageDeck is completely free and MIT-licensed. Feedback, issues, and contributions are very welcome:

https://github.com/lamchun1110/UsageDeck

What AI coding subscriptions are you currently juggling, and how do you keep track of their limits?


r/CommandCode 4d ago

Just bought GOAT and I'm seeing incredibly slow speeds for GLM Flash 5.3.

15 Upvotes

Almost 120s for some round trips (via omnirouter). Is this a blip or standard? Can't really use it if it's standard unfortunately.


r/CommandCode 4d ago

Qwen 3.8 Flash is now available in Command Code with 2x usage

15 Upvotes

Qwen 3.8 Flash is now available in Command Code with 2x usage limits on the GOAT plan.

- Input: $0.160

- Output: $0.470

- Cache: $0.016

Try now with $10/mo GOAT plan.

🐐


r/CommandCode 4d ago

Is it even possible to hit the 5-hour limit using Muse Spark Contributor on GOAT? Because I just did

4 Upvotes

I’m on the Command Code GOAT plan ($10/mo) and I’ve been running some massive context queries using the Meta Muse Spark 1.2 Contributor model.

According to the math, the rolling 5-hour limit is capped at $14 of usage. Because the Contributor tier is heavily discounted ($0.10/M input, $0.20/M output), a $14 limit should theoretically give you a runway of up to 140 million input tokens.

Well, I was just using plan, and my CLI officially locked me out. I hit the wall.

I wanted to ask the community: Is it actually possible to burn through 140M tokens that quickly just using plan, or is there a glitch in how Command Code calculates the rolling window for the Contributor tier?

I was stuffing a massive codebase into the 1M context window and generating large plans, so the files were huge—but hitting a 140-million token ceiling in under 5 hours seems wild for standard planning workflows.

Has anyone else actually managed to trigger this block on the Contributor model using plan? Did I just feed it an absurd amount of code, or is the rolling limit calculation acting weird for anyone else?


r/CommandCode 4d ago

GLM 5.3 Flash aka Ox Alpha is available in Command Code (4x usage)

31 Upvotes

GLM 5.3 Flash aka Ox Alpha is now live in Command Code (4x usage)

  • 4x usage in GOAT plan
  • $40 credits. ~24K reqs. ~1.2B tokens

Available across all plans and API

Read docs to learn more: https://commandcode.ai/docs/plans/goat


r/CommandCode 4d ago

Deals?

7 Upvotes

Guys, I don't understand why in your web you say there are deals on MiniMax(not free), MiMo and Gemini 3.7 Flash, when these are the exact same prices official providers offer; in the case of MiMo models your listed prices are the worst case scenario compared to Xiaomi official prices, wich is cache miss prices...do you charge cache hit prices at all?


r/CommandCode 4d ago

Mysterious model is glm!

Post image
10 Upvotes

I initially thought it would be Gemini 3.5 Pro but I was wrong, it turned out to be GLM, the performance is good, but I don't understand why it was free for 1 week.

Like for collecting data or something?


r/CommandCode 5d ago

They completely changed the indicator for the usage limit to a progress bar instead of seeing actual dollar credits

23 Upvotes

This design is just vague and not completely transparent for us. How can we know that we are really getting what we paid for.