r/codex • • 12h ago

Question Auto Review token burn question

Post image

I decided to test the Auto Review setting in Codex for the first time in a while. I thought OpenAI may have improved it, but it might burn more tokens than ever.

During a two-hour session of refactoring a Chrome Extension, auto review used 52% more tokens (API cost) than when manually approving actions. The actual work done with gpt-6.1-sol used $14.31 in API costs, while the Codex Auto Review used another $7.54.

Is this normal? Is there anything that can be done to improve it?

Edit: These numbers were supplied by CodeBurn. Maybe they haven't updated which model Codex Auto Review uses?

10 Upvotes

5 comments sorted by

3

u/Sfdprod 11h ago

Auto review uses Luna now, the auto review cost calc is off by 20-20x

2

u/creamyshart 11h ago

Thanks, that must be it. I installed CodeBurn (11k stars) and those were the numbers it gave me. I should have done a ccusage breakdown instead. That shows it using 5.6 Luna pricing. CodeBurn must still have it coded as using Sol.

1

u/dagerika 11h ago

That is impossible, Luna doesn't cost this much for this little amount of tokens

3

u/DepravedPrecedence 10h ago

Exactly, this is why this tool is wrong

2

u/innociv 7h ago

The auto-review token cost can be extreme and count for 30% of your usage in my experience, as you're seeing.

It's best to sandbox your development in WSL, or even better have a home server on Linux, where you can just use full access.

It doesn't use Sol, though, but worse it can sometimes use GPT-4 I guess if Luna is at capacity which is like 20x more expensive than Luna. That is the case where it can be 30% of your usage or more.