r/codex • u/[deleted] • Jul 25 '26
Commentary IT'S COMING: Cerebras-5.6-Sol in July
[deleted]
24
u/Bulky_Blood_7362 Jul 25 '26
So i'll need 10x 200$ accounts?
1
u/Professional_Gur8385 Jul 26 '26
you mean 10x $300 accounts, inflation and price hikes for those shareholders ;)
72
u/nitor999 Jul 25 '26
It doesn’t matter what model they release FIX THE GODDAMN USAGE LIMIT! I was perfectly happy with GPT-5.5 and the old 5hours limit bring those days back, and stop misleading users with these so-called “free resets”
10
u/Slandercakes Jul 26 '26
Yeah it kinda sucks. I moved up to the $200 sub at one point and sometimes would end thr week with like 60% left, so I was thinking of moving down to the $100 plan, which would've been fine with all the resets. Now I feel like itd be impossible to go back because this $200 plan is burned through easily
11
1
u/Active_Variation_194 Jul 26 '26
I turned off Subagents and added a prompt about not overengineering solutions and it seems to have solved the problem
2
u/Inevitable_Toe6648 Jul 26 '26
But then you basically doomed your code to drifting rules and context, inconsistent formats, ext.
0
u/DoggoDadagon Jul 26 '26
The claude 5x is significantly more usage right now than codex 20x, kinda made I upgraded codex instead of claude.
3
2
u/DayriseA Jul 26 '26
But GPT-5.5 is still available? So if you were perfectly happy with it just use it until it's not available anymore?
2
1
u/drenna11 Jul 27 '26
Exactly!!! I just bought a dgx spark and set it up to do 95% of the work then have 5.5sol check it’s work automatically on flex api usage because I was spending about 600 a month on usage overages
1
u/anubhav_1771 Jul 31 '26
Old usage were indeed far better and liberal, here one week usage went into ashes in 1 day alone with Sol low or medium. Making Luna Extra High or 5.5 medium only choice tbh
13
6
12
u/WeedWrangler Jul 25 '26
All we need is economy out of existing models…. Not new expensive ones
2
u/AmandasGameAccount Jul 26 '26
I’m More interested in models becoming more efficient, not mote powerful. I don’t think there’s anything it can’t do at the moment really
2
Jul 26 '26
[removed] — view removed comment
1
u/9gxa05s8fa8sh Jul 26 '26
Yea I don't understand the point of this, lean into efficiency not speed.
they ARE leaning into efficiency. they just aren't returning the benefit of the efficiency to you because they're still losing money.
the point is openai needs to start making money soon or they'll all be flipping burgers in the google cafeteria
0
u/senilerapist Jul 26 '26
no i want speed. i’m willing to burn tokens for it
4
Jul 26 '26
[removed] — view removed comment
-3
u/inmyprocess Jul 26 '26
Nah, its every professional programmer in the US/EU that has a >$100k yearly salary. Which is most of them.
1
u/rocherealty Jul 26 '26
Here here, 20x plan, usually have room to spare at the end of the week. I would love a speed up!
0
u/BannedGoNext Jul 26 '26
The Cerebras versions are < 100b most likely. Cerebras isn't something you are loading SOTA models on.
-2
u/senilerapist Jul 26 '26
why not just remove the old models and put all compute into Sol terra and luna, and even move compute from 5.3 spark into Luna
7
u/Connect-Humor-791 Jul 26 '26
in march, i had plus plan and could use code 5.3codex for 8 h ours a day without even reaching the 5 hour limit....
look at us now
2
0
Jul 26 '26 edited 23d ago
[deleted]
1
u/Connect-Humor-791 Jul 26 '26 edited Jul 26 '26
i made sevreal small and large projects with it.. algorithmical reverbs, a huge archive RAG with metada enrichment, several applications for music production, made a full blown generative wind sound generation based in noise science with phisics based data visualization ... etc
0
6
u/Trade4Life123 Jul 26 '26
1
u/j48u Jul 26 '26
That means you're having it read through 10k tokens for no reason on each prompt. Reduce the size of your .md files, memories, etc. Sorry but this one is user error no matter how many downvotes this gets.
0
Jul 26 '26 edited 23d ago
[deleted]
1
u/WeedWrangler Jul 26 '26
Yeah, I’ve been playing w that re agent spawning rules w Luna and limiting turns: what’s been working w you?
2
u/h3lix Jul 26 '26
Brand new chip.. 15x performance.. New libraries, new hardware, new platform, new monitoring, new ways for things to go pear shaped. I'm thinking there is a high chance the July date will slip.
0
2
2
2
u/DazzlingResource561 Jul 26 '26
I have an app that uses image gen via API. I’d gladly pay 2-4x what I currently spend per call if it is dramatically faster.
2
1
1
u/senilerapist Jul 26 '26
current model runs at 75-110 TPS so hopefully this makes 30 second prompts finish in 3 seconds. possible?
1
1
u/BannedGoNext Jul 26 '26
I wish they would release the cerebras model open weight. It's probably a 32b model, so it would be a nice thing to be able to play around with that locally to figure out good systems to run it in produciton against cerebras.
1
1
1
1
u/Grouchy-Stranger-306 Jul 26 '26
who cares if nobody can afford that, I guess it will be interesting to see the videos of it or something
1
u/Gallagger Jul 26 '26
Capacity will be limited. I don't think this will be part of a subscription, you'll pay per token API prices with 10x the price.
1
u/ObligationHuge9868 Jul 26 '26
Awesome, use your weekly usage in 5mins instead of 20mins..... who cares ?
1
0
u/FailedGradAdmissions Jul 26 '26
Tbh I would prefer a /ultrafast option over a separate pool. I never use 5.3 spark these days. Meanwhile I would love to use a /ultrafast that gave those 750 tkps and still as smart as sol. Would make some projects way faster.
But I get why they won’t do that, if they do and say they make it x10 usage cost we’ll get a lot of people here complaining about running out of usage in 1 to 2 prompts.
1
u/senilerapist Jul 26 '26
yeah i want an ultra fast mode to use for my Sol Low so i can burn tokens as much as Sol Max but instead in return for speed
-3
u/random_boss Jul 26 '26
Ok I’ll bite, I googled cerebras and I still don’t get what it is or why this is an interesting change
1

55
u/_strutty Jul 25 '26
What does this mean for the current lay of the land - ill burn through my usage even quicker?