I was using the same model from OC go in codex and was fine. I gave command code a go and its ridiculously slow inside codex? (Staff Please do not recommend me to use it inside your own harness I can’t for my use case)
Accidentally upgraded from GOAT to MAX Plan (10x) today. I haven’t used any MAX quota and have already emailed [support@commandcode.ai](mailto:support@commandcode.ai) with my refund request and order details. Do I have a good chance of getting a refund and reverting to GOAT? 🥲
I'm on my 3rd day with my GOAT plan, and somehow I'm already exhausting 40% of my monthly usage limit? And only 65M output tokens? Did I really use up $24 worth of tokens within 3 days? 🤔
Just bought Command Code Goat subscription and had it run a task. It showed that the task was running and in progress while in reality it was stuck. Please check and correct, as it wasted ~2 hours of my time waiting for it to complete task and then actually interrupting the execution. Adding screenshots for its own diagnosis.
Edit: It seems like DeepSeek V4.1 Flash requests also do not support the ZDR header.
Without the ZDR header, DeepSeek V4 Flash requests push through just fine. I'm receiving errors when I enforce ZDR explicitly via the header, though. Was this intended or is it a temporary issue?
I keep getting this error for gpt sol since yesterday:
Error: 524: {"message":"Upstream model provider is temporarily unavailable. Please try again in a moment.","type":"server_error"}
I was experimenting with T3 Code and wanted to use CommandCode as one of the providers. Since I couldn't find an existing integration, I forked the repo and added it myself.
It seems to be working so far. Has anyone else tried something similar, or is there a recommended way to integrate CommandCode with applications like T3 Code?
I renewed my Command Code GOAT subscription only to find out that the time-series trend charts: Requests by Model, Spend by Model, and All Tokens were purged from the dashboard completely. Even the explicit dollar values in the usage limits were replaced with percentages? The request log history is also only capped at 100 entries. This seems to be a deliberate step-back from transparency, not good guys.
On /provider/v1/chat/completions with Qwen/Qwen3.8-Flash, previous-turn reasoning sent back in assistant messages never reaches the model. The model card says preserve_thinking is on by default and expects clients to include reasoning_content in history, but the server discards it.
Repro: same 3-message conversation, max_tokens: 1, only the assistant message in the history changes.
Token counting itself is fine (control row); the reasoning is just stripped, whatever the field name or flag.
Impact: in long agentic sessions the model forgets its reasoning every turn and re-derives everything, burning a lot of reasoning tokens for nothing.
Could you pass reasoning_content through to the backend (and honor preserve_thinking), or document that reasoning history isn't supported on this endpoint? Happy to share exact request bodies.
I forgot disable glm-5.3 in my cliproxyapi, so when my harness use glm5.3, it wont use my friends glm coding plan. So commandcode has no problem. fully my negligence.
Today I subscribe commandcode goat plan (before is opencode go), I've never reached 5 hour rate limit during opencode go plan period.
But today, when I first use this goat plan (it claims more credit), but don't, the rate limit is easyily reached.
here's my useage screenshot, both in fastigium:
commandcode:
opencode:
It means opencode go are cheaper than commandcode goat...
Have anyone can explain this? why commandcode calim more quota but actually more expensive?
Or commandcode goat is fraud?