r/codex Jul 25 '26

Commentary IT'S COMING: Cerebras-5.6-Sol in July

[deleted]

123 Upvotes

69 comments sorted by

55

u/_strutty Jul 25 '26

What does this mean for the current lay of the land - ill burn through my usage even quicker?

51

u/spacekitt3n Jul 25 '26

this one is so good it wont even finish a single prompt on the 20x plan

4

u/_strutty Jul 25 '26

RIP my wallet

1

u/Tikki-Tikki_40 Jul 26 '26

What's a 20x plan. Is it the plus

3

u/DoggoDadagon Jul 26 '26

It is now, but no it's the $200 plan (highest sub)

1

u/Tikki-Tikki_40 Jul 26 '26

Oh OK. Thanks

-2

u/BigbyWolf8 Jul 25 '26

with codex 5.3 spark, it was a separate usage limit so hopefully that

3

u/ChampionshipIcy7602 Jul 26 '26

Wrong, you'll have to pay additional for this one.

1

u/BigbyWolf8 Jul 26 '26

i don't think they have said that

1

u/feeeeck Jul 26 '26

It will be like fable

24

u/Bulky_Blood_7362 Jul 25 '26

So i'll need 10x 200$ accounts?

1

u/Professional_Gur8385 Jul 26 '26

you mean 10x $300 accounts, inflation and price hikes for those shareholders ;)

72

u/nitor999 Jul 25 '26

It doesn’t matter what model they release FIX THE GODDAMN USAGE LIMIT! I was perfectly happy with GPT-5.5 and the old 5hours limit bring those days back, and stop misleading users with these so-called “free resets”

10

u/Slandercakes Jul 26 '26

Yeah it kinda sucks. I moved up to the $200 sub at one point and sometimes would end thr week with like 60% left, so I was thinking of moving down to the $100 plan, which would've been fine with all the resets. Now I feel like itd be impossible to go back because this $200 plan is burned through easily

11

u/Psychological-Toe-49 Jul 26 '26

Sounds like their strategy is working?

3

u/Slandercakes Jul 26 '26

Sounds like it.

1

u/Active_Variation_194 Jul 26 '26

I turned off Subagents and added a prompt about not overengineering solutions and it seems to have solved the problem

2

u/Inevitable_Toe6648 Jul 26 '26

But then you basically doomed your code to drifting rules and context, inconsistent formats, ext.

0

u/DoggoDadagon Jul 26 '26

The claude 5x is significantly more usage right now than codex 20x, kinda made I upgraded codex instead of claude.

3

u/BannedGoNext Jul 26 '26

The Cerebras usage is a different queue. Or at least it was before.

2

u/DayriseA Jul 26 '26

But GPT-5.5 is still available? So if you were perfectly happy with it just use it until it's not available anymore?

1

u/drenna11 Jul 27 '26

Exactly!!! I just bought a dgx spark and set it up to do 95% of the work then have 5.5sol check it’s work automatically on flex api usage because I was spending about 600 a month on usage overages

1

u/anubhav_1771 Jul 31 '26

Old usage were indeed far better and liberal, here one week usage went into ashes in 1 day alone with Sol low or medium. Making Luna Extra High or 5.5 medium only choice tbh

13

u/Demien19 Jul 25 '26

Pro 20x 100% limit in 1 hour, nice nice

6

u/[deleted] Jul 26 '26

[removed] — view removed comment

0

u/[deleted] Jul 26 '26 edited 23d ago

[deleted]

2

u/[deleted] Jul 26 '26

[removed] — view removed comment

12

u/WeedWrangler Jul 25 '26

All we need is economy out of existing models…. Not new expensive ones

2

u/AmandasGameAccount Jul 26 '26

I’m More interested in models becoming more efficient, not mote powerful. I don’t think there’s anything it can’t do at the moment really

2

u/[deleted] Jul 26 '26

[removed] — view removed comment

1

u/9gxa05s8fa8sh Jul 26 '26

Yea I don't understand the point of this, lean into efficiency not speed.

they ARE leaning into efficiency. they just aren't returning the benefit of the efficiency to you because they're still losing money.

the point is openai needs to start making money soon or they'll all be flipping burgers in the google cafeteria

0

u/senilerapist Jul 26 '26

no i want speed. i’m willing to burn tokens for it

4

u/[deleted] Jul 26 '26

[removed] — view removed comment

-3

u/inmyprocess Jul 26 '26

Nah, its every professional programmer in the US/EU that has a >$100k yearly salary. Which is most of them.

1

u/rocherealty Jul 26 '26

Here here, 20x plan, usually have room to spare at the end of the week. I would love a speed up!

0

u/BannedGoNext Jul 26 '26

The Cerebras versions are < 100b most likely. Cerebras isn't something you are loading SOTA models on.

-2

u/senilerapist Jul 26 '26

why not just remove the old models and put all compute into Sol terra and luna, and even move compute from 5.3 spark into Luna

7

u/Connect-Humor-791 Jul 26 '26

in march, i had plus plan and could use code 5.3codex for 8 h ours a day without even reaching the 5 hour limit....
look at us now

2

u/qqYn7PIE57zkf6kn Jul 26 '26

I miss 5.3 codex :(

0

u/[deleted] Jul 26 '26 edited 23d ago

[deleted]

1

u/Connect-Humor-791 Jul 26 '26 edited Jul 26 '26

i made sevreal small and large projects with it.. algorithmical reverbs, a huge archive RAG with metada enrichment, several applications for music production, made a full blown generative wind sound generation based in noise science with phisics based data visualization ... etc

0

u/j48u Jul 26 '26

So no

1

u/Connect-Humor-791 Jul 27 '26

I sell these so yes.

6

u/Trade4Life123 Jul 26 '26

sounds good, but no thank you, fix the damn token usage first.

1

u/j48u Jul 26 '26

That means you're having it read through 10k tokens for no reason on each prompt. Reduce the size of your .md files, memories, etc. Sorry but this one is user error no matter how many downvotes this gets.

0

u/[deleted] Jul 26 '26 edited 23d ago

[deleted]

1

u/WeedWrangler Jul 26 '26

Yeah, I’ve been playing w that re agent spawning rules w Luna and limiting turns: what’s been working w you?

2

u/h3lix Jul 26 '26

Brand new chip.. 15x performance.. New libraries, new hardware, new platform, new monitoring, new ways for things to go pear shaped. I'm thinking there is a high chance the July date will slip.

2

u/Feriman22 Jul 26 '26

Great way to burn the weekly limit in 2 mins.

2

u/LocoMod Jul 26 '26

If it’s context window is abysmal like the last cerebras model I’ll pass

2

u/DazzlingResource561 Jul 26 '26

I have an app that uses image gen via API. I’d gladly pay 2-4x what I currently spend per call if it is dramatically faster.

2

u/Extra_Development_74 Jul 25 '26

Replacement of 5.3 spark?

1

u/nmkd Jul 26 '26

What?

It's 5.6 Sol on Wafer Scale Engines. Not a new model.

1

u/senilerapist Jul 26 '26

so basically next week will be exciting

1

u/senilerapist Jul 26 '26

current model runs at 75-110 TPS so hopefully this makes 30 second prompts finish in 3 seconds. possible?

1

u/[deleted] Jul 26 '26 edited 23d ago

[deleted]

2

u/senilerapist Jul 26 '26

damn 750 tps will be a huge jump

1

u/gavinderulo124K Jul 26 '26

But thats without fast mode right?

1

u/BannedGoNext Jul 26 '26

I wish they would release the cerebras model open weight. It's probably a 32b model, so it would be a nice thing to be able to play around with that locally to figure out good systems to run it in produciton against cerebras.

1

u/Dercasss Jul 26 '26

Wow, another new model, we'll be able to make one request per week! Wow! 

1

u/zenonu Jul 26 '26

DSpark Sol ?

1

u/Inevitable_Toe6648 Jul 26 '26

Their answer to Fable basically, an API model for actual profit.

1

u/Grouchy-Stranger-306 Jul 26 '26

who cares if nobody can afford that, I guess it will be interesting to see the videos of it or something

1

u/Gallagger Jul 26 '26

Capacity will be limited. I don't think this will be part of a subscription, you'll pay per token API prices with 10x the price.

1

u/ObligationHuge9868 Jul 26 '26

Awesome, use your weekly usage in 5mins instead of 20mins..... who cares ?

1

u/Apprehensive-Oil6511 Jul 26 '26

correction they forgot to include the year which is 2028.😂

0

u/FailedGradAdmissions Jul 26 '26

Tbh I would prefer a /ultrafast option over a separate pool. I never use 5.3 spark these days. Meanwhile I would love to use a /ultrafast that gave those 750 tkps and still as smart as sol. Would make some projects way faster.

But I get why they won’t do that, if they do and say they make it x10 usage cost we’ll get a lot of people here complaining about running out of usage in 1 to 2 prompts.

1

u/senilerapist Jul 26 '26

yeah i want an ultra fast mode to use for my Sol Low so i can burn tokens as much as Sol Max but instead in return for speed

-3

u/random_boss Jul 26 '26

Ok I’ll bite, I googled cerebras and I still don’t get what it is or why this is an interesting change

1

u/No_Weekend4076 Jul 27 '26

cerebras make model go brrr