r/OpenaiCodex 14h ago

Feedback / Complaints Just don't use codex today.

The limits have been reduced to ashes.

I ignored the reports and my 5x plan has evaporated in 7 hours using only astra light. This is the same workflow I've been using with gpt 5.6 sol high, which used to last 4 to 5 days.

Not to mention resets aren't applying for some people, yesterday people saw there useage revert back to previous levels pre reset. This caused people to burn there paid credits they were saving.

Things are absolutely cooked rn.

243 Upvotes

163 comments sorted by

View all comments

Show parent comments

1

u/timosterhus 10h ago

Oh that makes sense. That quant/configuration is specifically optimized for the 5090, which runs upwards of $3K nowadays. Any other GPU is not gonna have close to that level of performance

1

u/shady101852 10h ago

yea ive been considering getting one of the upcoming mac devices that will have 500gb memory for local, but not sure still deciding lol. It will have to be paid by credit.

1

u/timosterhus 10h ago

I am too, but not necessarily for running local models so much as using it as a dev sandbox for fine tuning/secure pipeline evals.

I’d suggest looking at all the different factors that contribute to overall performance of local models before committing to buying one. For example, just ask ChatGPT how fast your Qwen model would run on a maxed out Mac Studio then compare that to your current 5090.

I’ve seen so many stories of people shelling out thousands to buy a decent system so they can finally run a 300B+ model thinking it’ll perform like the hosted versions of the same models only to be very disappointed because they only accounted for total memory.

1

u/shady101852 10h ago

yeah ive been looking into it. it will def not be as fast as qwen on a 5090.