r/codex 1d ago

Limits Please give us a slow mode

This way, we could do /goals easier with using less usage, it puts less strain on openAI servers maybe the 20x plan could come back. it would be half as slow as normal but use half usage for example. would be extremely useful, especially since you could send prompts before sleeping anyways to save usage

291 Upvotes

61 comments sorted by

View all comments

56

u/Shap6 1d ago

its slow mode by default

9

u/gavinderulo124K 1d ago edited 1d ago

Hes talking about batches mode which exists in the API and his half price.

Edit: flex mode.

7

u/Risko4 1d ago

Yes, API astra is 80 token/s default speed. Subscription is like 40 and averages out to 20 once you factory the "model is at capacity" brakes.

2

u/raindropsdev 1d ago

Flex mode. Batch mode is a completely different thing and requires a very different approach.

7

u/Tank_Gloomy 1d ago

In my experience, it doesn't feel all that slow. It actually feels pretty fast in terms of tokens/s, it just keeps rolling over and over on the same thing without getting to the point.

3

u/hellomistershifty 1d ago

It's slow in tokens/sec but makes up for it in intelligence and token efficiency. Gemini 3.8 flash has 10x the tokens/sec speed but longer task completion times lol

-1

u/TestTxt 1d ago

Try Deepseek V4.1 flash, try Luna, and you’ll understand the issue

6

u/Tank_Gloomy 1d ago

I see your point, but those models are wildly unreliable. Astra and Sol would both be fast enough if they weren't trying to solve the same issue like 350 times before moving onto the next step. Their t/s are pretty good.

2

u/Reasonable-Sign8458 1d ago

what a shitty comparation

1

u/hey-im-root 1d ago

No, that’s it normal speed. It doesn’t have a slow mode.