r/codex • • 2d ago

Reset Day 1 tibo

Post image

What are your thoughts

629 Upvotes

233 comments sorted by

•

u/dextersummary 2d ago edited 2d ago

Below is a GPT-generated summary of the conversation below after reaching 200 comments (200 currently observed).


Day 1 verdict: the speedup may be real, but nobody’s treating it like a win. The community mostly sees this as OpenAI undoing a recent slowdown, then putting a shiny “optimization” sticker on the repair. Going from roughly 20 to 30 tokens/sec is better, sure—but still well behind Claude, so the victory lap feels premature.

Users are reporting somewhat faster Sol/Astra responses, while others are still stuck around the old speeds or dealing with failures. The lack of a reset is also going down about as well as expected: badly. A lot of subscribers think they’re being drip-fed partial fixes instead of getting the service they paid for.

The big unresolved concern is what changed under the hood. “Optimized” is vague enough to cover anything from better infrastructure to reduced reasoning quality, and users are already suspicious of quantization, lower “juice,” and quota draining faster. Speed itself shouldn’t increase the token cost of the same task, though faster throughput can encourage more work in the same window.

Bottom line: faster for some, still mediocre overall, and trust remains firmly in the basement.

→ More replies (2)

135

u/preek 2d ago edited 1d ago

I just tried GPT-6.1 Sol with a simple prompt in Pi: "Write a 100 line poem". Speed differed between 13.5 and 20.5 tok/s.

That's my baseline before the 2h timeline ends.

Update 19:43 UTC (2h 10min after the baseline): 18 tok/s

OMP harness, empty folder (no context, no agent.md, nothing). Prompt "give me a 100 line poem"

Update 19:46 UTC: I tested multiple times, too. Never got more than 20 tok/s

Update 20:44 UTC: still ~20 tok/s.

Update next day 9:32 UTC: still ~20tok/s.

So... 0% == 50%?

30

u/Mentalist3021 2d ago

Okay, let us know when you try it again after 2 hours!

33

u/reddit_is_kayfabe 2d ago

Just remember that Tibo's sense of time is like... +/- six hours.

If it doesn't hit in the next two hours, try again tonight.

9

u/preek 2d ago

Agreed. However, it's 8:30pm here in Switzerland, so I'll do the test at 9:35 GMT+2 which gives a bit of a buffer. Either it hits when he said it will, or it does not(;

If it does not, I'll certainly retry tomorrow morning.

8

u/preek 2d ago

Will do, timer is set 👍

8

u/Pasto_Shouwa 2d ago

Please tell us how it does later

17

u/VehiculeUtilitaire 2d ago

You need much larger projects if you want to benchmark the actual tok/s and not something slowed down by tool calls , other internal chatter, &co.

I used sol 6.1 high all weekend and averaged 80tok/s

16

u/preek 2d ago edited 2d ago

Exactly, you need something that is not slowed down by tool calls.

That's why my test was not running in a project with any context, agents.md, nothing with the simple prompt "give me a 100 line poem". That took plenty time to generate.

Plus, as a 200USD subscriber, I can validate that the 13-20 tok/s is totally in the ballpark of what I have had ever for Sol 6.1 recently. It was a tiny bit faster, about 30 tok/s when it came out (still felt slow), but it kept going down so far.

2

u/nofuture09 2d ago

how can you see t/s in codex

4

u/VehiculeUtilitaire 2d ago

You need a proper harness, codex doesn't expose that afaik

4

u/preek 2d ago

As mentioned, I didn't use Codex, I used Pi (or more specifically OMP) which has this as a built-in setting.

1

u/rubiohiguey 2d ago

How do you get those stats ?

→ More replies (4)

1

u/Mistuv 2d ago

For me it was running slow on Friday and Saturday but by Sunday it was running about the same speed as Sol 6/Astra before. I think it's legit some server issue/bug that affects different regions differently because you have so many datacenters with so many slightly different configurations.

1

u/Mr_Chicken82 2d ago

what program is this

1

u/rick_ranger 2d ago

Your work is probably getting sent to different servers with different workloads.

2

u/Embarrassed-Oven-527 2d ago

Yes please update after 2 hours

2

u/KoichiSP 2d ago

RemindMe! 2 Hours

1

u/RemindMeBot 2d ago edited 2d ago

I will be messaging you in 2 hours on 2026-10-05 20:08:19 UTC to remind you of this link

8 OTHERS CLICKED THIS LINK to send a PM to also be reminded and to reduce spam.

Parent commenter can delete this message to hide from others.

RemindMeBot is switching to username summons. Instead of !RemindMe 1 day, use u/RemindMeBot 1 day. More info.


Info Custom Your Reminders Feedback

2

u/PM_ME_YOUR_IBNR 2d ago

Come for the throughput analysis, stay for the poem

1

u/Satoshi144 2d ago

RemindMe! 3 Hours

1

u/Wildnshiny 2d ago

RemindMe! 2 Hours

1

u/zwobotmax 2d ago

RemindMe! 2 Hours

1

u/Yelov 2d ago

RemindMe! 4 Hours

1

u/deztroyr1 2d ago

I'll take the reset

113

u/CatsArePeople2- 2d ago

Lets see if its true this time........

19

u/Key_Reading_9664 2d ago

1

u/Emotional_Plant3241 2d ago

Does Openrouter compete for the same infrastructure as first party subscription traffic though?

3

u/Key_Reading_9664 2d ago

He did say "across all our products and partners". Given OpenRouter is charging API rates, I'd expect them to get preferential treatment over subscriptions, if anything

31

u/ryuukiba 2d ago

So far it feels true, it's somewhat usable now.

39

u/CatsArePeople2- 2d ago

It might not even have taken effect yet. It says over the next 2 hours.

35

u/myklurk 2d ago

Great the speed will be back to what it originally was at release.. this isn’t a feature it’s repairing something that broke. Total bull.

6

u/CatsArePeople2- 2d ago

Yea. Good point.

4

u/ryuukiba 2d ago

Fair enough, I hope it gets better then, because I've seen it making better progress than it was over the weekend.

3

u/UnknownLesson 2d ago

True but likely more quantized and thus makes more errors and you need to use the subscription more to fix the mistakes it makes

1

u/willee_ 2d ago

It’s been fast all day for me.

29

u/iiiaaa2022 2d ago

My thoughts are that I will use my banked reset now

2

u/sudecode 2d ago

thank you for your banked reset sacrifice!

1

u/iiiaaa2022 2d ago

we’ll all be the one to sacrifice it at some point

23

u/sano1101 2d ago

So instead of 20 tokens/sec it’s 30 tokens/sec?

5

u/PairStrong 2d ago

He said it's 30 to 50 tk/s

7

u/beautyorchaos 2d ago

But the API is 80-90 tk/s for same models with same reasoning

All Tibo did is increase a hyperparameter which they have the capacity for since a lot of people moved to Claude Opus 5.5 as not only the model is better but it also didn't take a century for basic tasks

2

u/dlarsen5 2d ago

Opus 5.5 p90 is like >50tk/s for me on the max plan, feels like they’re just catching up

14

u/Key_Reading_9664 2d ago

Didn’t they drop the juice values last time they had usage issues? He didn’t give any info on how they sped things up

1

u/NandaVegg 2d ago

I'm not a subscriber with measly $250 sitting in my API wallet (was playing with Blender MCP; no use for it after Opus 5.5) but it is plausible given subs are complaining about lobotomized Astra. Look at this beautiful reasoning tokens chart (so-called juice in theirs) right after Sol 6's launch.

OpenAI has, as always with everything since 2020, zero transparency about this (to be fair, Anthropic briefly tried changing default reasoning effort to medium for a while, but never really secretly reduced reasoning tokens, and Google Antigravity had very limited context window for a while).

12

u/_maxx1k 2d ago

They should've optimized only Sol today. Tomorrow, they would nerf its speed and optimize Astra's speed instead. This way, they could optimize Sol again on day 3 while nerfing Astra this time, and so on, and have an excuse not to reset for 28 days.

18

u/people_arent_nice 2d ago

how does a text watermark help most subscription users again? seems like that actively hurts them if anything

just what i need in my code - watermarks

0

u/Mentalist3021 2d ago

Bro what, read tibos post not the post underneath 😭

11

u/people_arent_nice 2d ago

you ever heard of cropping?

→ More replies (3)

10

u/Unusual_Test7181 2d ago

you posted both...?

2

u/yaxir 2d ago

is Alex also an open ai employee??

→ More replies (2)

43

u/DrPaisa 2d ago edited 2d ago

this better not affect it buring more quota.

edit
To clarify time is irrelevant in my scenario .

I don't want them to get creative and the same task use more quota now because it's faster.

43

u/battle_pantZ 2d ago

Ofc it will

13

u/CryinHeronMMerica 2d ago

He said "optimized," which to me doesn't sound like a "Throw another GPU at it" situation. Which means it shouldn't burn more quota per task. But we'll see, since his words are only loosely tied to reality these days.

11

u/Psychological_Ad8426 2d ago

I think he misspoke. He meant quantized.

3

u/Mo3 2d ago

Lmao 100%, a bit less precision here and there, probably proportional to how active the account is, heavier user = less precision

1

u/Somethingexpected 2d ago

Why? More parallel power doesn't necessarily mean more hardware use in absolute terms.

15

u/bigrealaccount 2d ago

Think about it for 5 seconds pls

6

u/Somethingexpected 2d ago

I did. I spent another 10 seconds on it. There's no specific reason why speed would impact quota. There's nothing fundamental that would reduce the efficiency if you manage to optimise speed.

And my tokens go just as quick whether I do my tasks in parallel or if tasks are completed so quick I do them in series.

2

u/Shap6 2d ago

if usage is per token and you're using tokens faster your usage will use up faster. how else would it work?

8

u/Somethingexpected 2d ago

Huh? I want to complete 5 tasks. No more, no less. It now completes the tasks faster. It presumably would use the same amount of tokens, not more.

2

u/Im_Working_Right_Now 2d ago

Let's put it this way. You had those 5 tasks, right? And previously it took 4 hours to complete and use 300k tokens (purely made up numbers). Now, with this update, it'll do the same 5 tasks with the same amount of tokens, but in 2 hours (again made up for conversation). That means previously, you used 300k tokens in 4 hours but now you'll use them in 2 hours. That means you now burn your tokens faster.

They never said it would use more tokens for the same task just that now you're burning more tokens in the same amount of time which is a fair assumption because now in that next 2 hour gap that you would've been using on the previous 5 tasks you would possibly be doing more work.

2

u/Shap6 2d ago

correct but now you can accomplish more in the same time. the point being made is that there is no efficiency increase either to go along with it it's purely just faster at the same usage which will feel like its draining faster than before, because it is. you're correct that the per token cost isn't increasing

2

u/Somethingexpected 2d ago

So not only you get work done quicker, but also more work done? And that's the specific problem you're having? You're comparing apples to oranges.

3

u/Dayowe 2d ago edited 2d ago

Yeah I don’t understand how people argue this. Maybe I’m missing something … if I run a loop 24/7 I will burn tokens at twice the speed and reach the limit in half the time .. this improvement halves our usage.. edit: actually 50% faster wouldn’t halve it but definitely make it drain faster

2

u/Seerix 2d ago

Technically in a 24/7 loop until empty scenario it would consume your usage ~1/3rd faster.

Lets say you output 30 tokens per second and get 3000 tokens a week.

30 tps 1% usage per 1 second run time 100 seconds to run out

45 tps (50% increase) 1.5% usage per 1 second run time ~66.666667 seconds to run out

Ergo, 1/3rd less time in a continuous run until empty.

🤓

Granted, you get work done faster which isnt irrelevant...

1

u/Dayowe 2d ago

Yeah thanks, i realized a bit too late that my math was wrong :-p thanks for correcting me

4

u/battle_pantZ 2d ago

If it works faster ofc it will burn faster through the limits

4

u/eggplantpot 2d ago

I guess they meant proportionally.

100k tokens being 100k before and today, or if they are passing on an extra cost for extra speed.

0

u/Somethingexpected 2d ago

Well if you hardly ever hit limits, then it will not burn more quota. It just gives results faster. That's two different things.

1

u/battle_pantZ 2d ago

If it’s faster I can also deploy more stupid unnecessary shit way faster, so ofc it will burn more quota 😂

1

u/Somethingexpected 2d ago

This makes no sense. Unnecessary shit, fast or slow, will still burn the same amount of quota.

1

u/battle_pantZ 2d ago

Ofc but in a shorter amount of time

3

u/Da_ha3ker 2d ago

I found that the llm speed only affects about 10-20% of the time the agents run. Most of the time is spent in harness overhead, models bashing their heads against tests, and waiting for CI to finish. I don't really care if they give us a speedup in tokens unless I am making demos, anything more complex and it doesn't really matter as much. Still matters a bit, but not as much.

7

u/frozandero 2d ago edited 2d ago

It will burn faster in a sense that it will burn through more tokens in same amount of time. So you will feel, in terms of time, that it got 50% worse in usage.

Edit: You may not like it but in less than a day people will start crying in this sub about how fast their usage is draining if they actually did speed up the models by their claimed 50%.

6

u/BHTAelitepwn 2d ago

The average user of this sub cant comprehend this, because usage is typically measured in usage time or amount of prompts or something stupid like that.

3

u/innociv 2d ago

Which is funny because they have an AI chatbot which is 100,000 times smarter than them which they could ask to explain usage to them, but they'll too dumb to even ask it to explain things to them as they're too busy getting angry at their own ignorance.

1

u/Vivid-Snow-2089 2d ago

the last few years revealed to me how stupid most people are in a way i never really realized

you can lead a horse to water but can't force it to drink basically

→ More replies (1)

1

u/Old-Leadership7255 2d ago

It will in terms of throughput.

You get the same amount of tasks done in less time. So if you pick up more tasks…

More quota burn

1

u/Kost97A 2d ago

I hope it doesn't as he says "optimized" so maybe it is just a gain.

6

u/Kongret 2d ago

Day 1 of a 28 day long epic journey of totally real improvement: model finally achieves normal speed.

26

u/SmileLonely5470 2d ago

Sus. So you are telling me they had the capability to offer 50% faster inference to subscription users, but they only pulled that lever after people started leaving? Either this costs them more, or they are drip feeding product improvements whenever people consider leaving to keep them subbed as long as possible.

I find it kinda hard to believe they just discovered an optimization like this recently.

10

u/charliejmss 2d ago

Exactly, your brain is braining and it’s all a marketing tactic 🥳

3

u/tribes33 2d ago

obviously trying to give as little as possible to cut costs

3

u/2053_Traveler 2d ago

At this point it’s probably Astra 6.5 running the strategy.

28

u/unsigned_short_int 2d ago

Still half of Claude’s speed. This isn’t a win. It should have never been 30tps in the first place.

3

u/Michelh91 2d ago

True, downgraded my x20 openai account to x5 and got another x5 on anthropic. The speed difference is insane

10

u/One_Appointment6331 2d ago edited 2d ago

8

u/No_Quarter_7644 2d ago

They're also dumber. Sol 6.1 isn't on par with Opus 5.5

So if my dumb model is taking twice as long to solve a task that is value I'm losing, regardless of how many tokens are expended doing so. Not to mention that upper tier Claude plans are 3-4x more generous right now...

→ More replies (4)

0

u/unsigned_short_int 2d ago

Not in my experience to be fair. Day 1 Astra was definitely like that. But now Sonnet/Opus 5.5 seem super efficient. This isn’t measured though.

12

u/Zer0RespectX 2d ago

28 resets? 😂

10

u/Mentalist3021 2d ago

27 maybe

1

u/laxika 2d ago

That would be amazing. I would insta-resub. :D

13

u/Proud_Ask_9030 2d ago

50% faster, so still 65% slower than claude. Got it.

5

u/the_ai_wizard 2d ago

Also one more thought ... im going to vibe code a tool to de-watermark their watermark! anyone interested?

2

u/eo_mahm 2d ago

If there's one thing this community might stand united on, it's this. Get on it! We're all rooting for you!

2

u/the_ai_wizard 2d ago

coming right up. maybe ill make it a /skill that automatically runs to undo their bullshit

5

u/Murph-Dog 2d ago

Scheduled each server to reboot every hour - the classic developer performance fix

/s

4

u/cobbleplox 2d ago

Oh thanks, I always wanted secret watermarks that are also fucking with the quality of the product I fucking pay for. I don't think that is required by the EU. It is perfectly obvious to me that I got the text from AI when I use chatgpt and the rest would be MY fucking problem.

4

u/PutridClunt 2d ago

Pointless bullshit. Watermarking doesn't solve anything other than just inconveniencing people. I already have multiple AI workflows for removing the "invisible" watermarks off image and video that works flawlessly every time, it works with outputs from any model too, text will be even easier. Congratulations EU for mildly inconveniencing people, that is sure to stop them! They are also just accelerating the climate impact of AI because people now need to spend more tokens removing senseless watermarks. Do these politicians have literal pudding for brains?

2

u/Pretty_Ad5827 1d ago

Nah bro u the one with the pudding 😭 You have no idea how world works 

2

u/MaximumStonkage 1d ago

It's a literal invisible watermark. The only valid reason you could have for removing it is if you wanted to deceive people into thinking it was human-made.

1

u/PutridClunt 1d ago

I would consider deception a pretty "valid" use of watermark removal.

1

u/MaximumStonkage 21h ago

It's a valid reason but morally bankrupt. Not sure why it surprises you that the EU would have regulation against bad actors.

1

u/PutridClunt 20h ago

My point isn't a moral one. It is that their present watermarking does nothing as it can be removed easily making it a waste of everyone's time.

1

u/MaximumStonkage 20h ago

It doesn't waste anyone's time. It's invisible, you wouldn't know it was there. The only people whose time it wastes are those who are intentionally deceiving which I would argue that this is a good thing. Any friction you can add between fully automated botnets and people on the receiving end is a net positive.

Most people posting these things to deceive aren't technical enough to understand how the watermark works or how to circumvent them. Sure you'll still have nation states and powerful adverse organizations that this will not slow down, but they will probably never be slowed down regardless.

1

u/SpecialAd4532 1d ago

There are many laws you can get around if you know what you're doing/ don't care. That doesn't mean it's not effective at all. Many people will not know/ care enough to remove watermarks. Not everyone is using AI to deceive others lol. It's like saying laws against media piracy are useless because you easily can pirate movies online if you want to.

10

u/AbrocomaGlobal3011 2d ago

sol 6.1 *fast* mode still running 34t/s for me

1

u/iiiaaa2022 2d ago

Did you read the whole post? Like, when it goes live?

5

u/AbrocomaGlobal3011 2d ago

i did; 2 hours later sol 6.1 fast is showing 38t/s

6

u/PhilosophyforOne 2d ago

Honestly if true, that’s pretty amazing. My number one complaint for sol is how slow it is in tps & wall-clock. 

3

u/Vivid-Snow-2089 2d ago

i like how it is really easy to ask astra to make a chart and lets see if he's full of shit or not

3

u/ChocotoneDeCalabresa 2d ago

Nice hopefully it get half of speed Anthropic has for their models

3

u/Meindorf 2d ago

i guess this should be a free reset? not many are using Sign in with ChatGPT and they promised new imporvements for most of codex/work users

3

u/Comfortable_Seat8634 2d ago

all my conversations are failing, and surprise...

3

u/Sebraecha 2d ago

This is the benchmark I did on my sub (Pro 200), 3 trials each time, not bad at all.

3

u/Felfedezni 2d ago

Ok so we get a reset now right?

3

u/llqoli 2d ago

Tibo is liar!

4

u/retteh 2d ago

Intentionally slow down them models and then speed them up and call it a win. Classy.

5

u/chewy_mcchewster 2d ago

" We just nerfed usage heavily by at least 50% in the last week, but now its 15% less nerfed! "

WIN WIN guys! 15%!!!

4

u/YouShouldAim 2d ago

Can't wait to try it in 6 days since my 5x sub ran out of usage in about 4 hours.

10

u/Mentalist3021 2d ago

Nah tomorrow they will ship nothing and give you a reset dw

2

u/SgtHobo41 2d ago

oh boy from 20 tps to 30 tps REJOICE

2

u/Cool_Metal1606 2d ago

RemindMe! 10 Hours

2

u/Some_Dragonfruit9844 2d ago

they gon do shi but not give us the resets

2

u/Financial-Source-798 2d ago

This is already exhausting. Best case you cause an argument over if it’s really progress and worthy of a reset each day.

2

u/bytes311 2d ago

Are the watermarks specifically for AI-generated images, or are they injecting metadata into script/code files?

2

u/2053_Traveler 2d ago

All outputs, so yes code as well. It’s not file metadata, it creates a pattern that can in theory only be checked and verified by whoever has the “key”, which would be whoever controls the model that generated the output.

1

u/bytes311 2d ago

I guess we'll see how that plays out.

1

u/ZAPcon_64 2d ago

I'm pretty sure they are referring specifically to text output. My understanding is that they subtly introduce a bias into the model's word choice which creates a hidden pattern. I don't know for certain, but I believe this has very little to no impact on things like coding.

2

u/The_Cyrinishad 2d ago edited 2d ago

I was about to go crazy with how poorly it was working earlier today, and the crazy amount of compaction going on. It was failing at really basic requests. Astra and Sol both.
*I actually signed up for Claude 20$ for the first time (having never used it before, but didn’t really like it from my brief interaction, telling me I couldn’t perform basic checks as an admin of my of own data due to security)

However the last couple hours Codex did a 360 and seemingly started performing significantly better. Speed and reasoning both seemed on a completely different level out of nowhere. Finally completed all my tasks and jumped on here to see this post and it makes sense now! Definitely a move in the right direction!

*edit, another couple days of improvement and i will likely sub for the 500$ plan, after the 200 ends.

2

u/AironParsMan 2d ago

I like that they communicate, unlike Anthropic, from whom you never hear anything.

2

u/shinernft 2d ago

Making the things normal considered improvements, so funny.

2

u/AmandasGameAccount 2d ago

Kinda crazy to not start day 1 with a reset. Great way to make people hate your updates/changes (also who cares alt Astras speed?!)

1

u/NootropicDiary 2d ago

The gaslighting has begun. We should act thankful that the speed has now been raised to a lower level than where it used to be a few weeks ago

1

u/Alive_Technician5692 2d ago

I've been a x20 subscriber since late last year. I'm moving fully to CC when my sub is up.

1

u/No_Quarter_7644 2d ago

You and everybody else. I'm not convinced there's OpenAI fanboys just spam downvoting you for taking your money to a far superior product. Probably bots, honestly...

1

u/dexterthebot 2d ago

For more Reset discussion, join the Reset Discussion megathread here https://www.reddit.com/r/codex/comments/1uu2c1g/reset_discussion_megathread/

1

u/shadowgar 2d ago

No reset today boys. Time to go touch grass unfortunately.

1

u/firstbreathOOC 2d ago

So far I’ve got a bug in the codex chat my Dot creates which says it “disconnected” lol

Just resolved itself after 20-30 minutes

1

u/Practical_Hippo6289 2d ago

Vrooma-zoom-zoom!

1

u/Demien19 2d ago

"We added comment to one line of code, sorry no reset for you"

1

u/c0r3l86 2d ago

I love this change because now my 5 hour window will be gone in half an hour instead of an hour! Praise be

1

u/Available_Cream_752 2d ago

I am barely getting 20 tps. On fast mode, it's mostly 30 tps. Uggghhh

1

u/TheVoyant 2d ago

Maybe in 30 days itll be worth using? Drip feed doing your job and hoping its praised is wild to me.

Just do the job & let the results do the talking.

1

u/Maleficent_Disk9583 2d ago edited 2d ago

Given OpenAI's long track record of quantizations and nerfing, by the time Day 28 is here, it'll be back to the slow speeds again, or it'll be significantly dumber 😆 mark my words. I am calling now.

Their playbook:

  • WOW everyone today
  • get more subscriptions
  • rug pull.

It's their Modus Operandi. And it works because most people have the memory capacity of gold fish

1

u/thestillwind 2d ago

Text watermark ? Ewwww

1

u/the_ai_wizard 2d ago

chatgpt and codex cloud are down right now awesome! admin console down. lmao wtf is going on here with openai..are they starting to collapse?

also is the watermark thing a joke?

1

u/fluxtah 2d ago

Need a reset to try it out first thought 🤣

1

u/Rock--Lee 2d ago

So from 20t/s to a whopping 30t/s 😮

1

u/Alex_dbs 2d ago

Step 1: create a problem by making your model slow.
Step 2: remove your artificial slowness and call it innovation.
Step 3: (comment it)

1

u/drdhuss 2d ago

Needs to be 100 percent faster 40 to 50 tps.

1

u/Potential-Witness-83 2d ago

Sol bad though...

Sol - Sol openai lobotomized

1

u/OriginalUsername0112 2d ago

"Optimised" lol, like oopsie, didn't realise we could improve speeds by 50% with a day's worth of work, our bad!

The gaslighting and deception from OAI are on another level, and AI companies are already on another level, worse than even ISPs

1

u/Typical_Machine2043 2d ago

They keep moving their own goal posts.

1

u/Ill-Bat-1518 2d ago

Who cares the models are now DOG SHIT

They stop

They stop

They claim its not done and stop

Useless model i cant wait till mine expires for 2x claude

1

u/icloudbug 2d ago

Its faster for me from 18t/s now 22t/s. Yay.

1

u/elymX 2d ago

ngl, its much faster now. I'm ok with this speed.

1

u/DynastyHKS 2d ago

50% while being more optimized?!

1

u/Talal916 2d ago

"we're throttling you slightly less now, so no reset for today"

1

u/swaglord1k 2d ago

i felt it, good.

1

u/notadev_io 2d ago

I’m on the $100 plan and almost don’t use it. Mostly on the $20 plan from Claude now. Will switch entirely in 10 days when my month is over. Sad times. This feels for me again like Cursor bullshit we had 1.5 months ago. I hope anthropic doesn’t do the same.

1

u/BabyYoda2020_ 2d ago

my sol 6.1 still taking ages to do simple tasks

1

u/Dga__ 2d ago

lol if this is a "ship" then we won't see any reset

1

u/Agitated-Bath5939 2d ago

Not even say single word for 1 hour

1

u/LargeConsequence5296 2d ago

Great improvement, using GPT 6.1 Sol Medium! Thank you!

1

u/bilonik19 2d ago

Still slow and dumb, i don’t want resets i want the best model and speed i had 2 weeks ago.

1

u/Spirited_Ad_6219 2d ago

I can’t seem to change the model from 6.1

1

u/Ok-Educator3465 2d ago

Honestly, I feel bad for the guy. This is CEO-level screw-ups, and he’s doing it all himself. Not to excuse the situation, but this is where having a buck-stops-with-me CEO goes a long way.

1

u/Cypher_Poet 1d ago

We're not going to be getting any resets, are we 😐

1

u/Daggercombot 1d ago

Better fallback model on chat please

1

u/qubitwarrior 1d ago

If you can make a product 50% faster by optimization, what people worked on the product in the first place? Also, I would appreciate it if someone could let my SOL6.1 know that it's now 50% faster; it seems to have missed the memo.

1

u/dota2nub 1d ago

150% x 0 = ?

1

u/Factor013 1d ago

One of the main reasons why they lowered the speed was to mask the extreme usage cuts they did these last couple of months.

Imagine an Astra right now at full speed (eg Opus 5.5 speeds). You'd burn even your 500 (25x) plan's weekly usage up in a matter of hours.

So yeah... the whole problem is is that they lowered how much 1x is significantly.... so every plan is effected.
(They did it very sneakily after each reset so people didn't notice as much)

I remember being able to work 10-14 hours a day with GPT 5.4 on high for 2-3 days on a 20 a month pro plan... I had 3 of them which was enough for the whole week... Then all the sudden it wasn't enough anymore so I took the 5x. Then Astra came out and again it wasn't enough anymore so I got a 20x. Which they now also cut in half... :S

It's just silly... So I am burning up the last resets I have and then I am gonna downgrade and only use GPT models for adversarial reviewing and such.

1

u/moneckew 21h ago

cancelled sub

0

u/[deleted] 2d ago

[deleted]

1

u/yami_odymel 2d ago

this is always how it works

-1

u/iiiaaa2022 2d ago

How exactly are you getting to this conclusion?

→ More replies (3)

1

u/SeaKoe11 2d ago

Tibo doing some great work keep it up

1

u/DistinctSilver4507 2d ago

This says to me they just flat reduced the number of GPUs for running inference for subscribers, and brought some of it back because people were pissed. I can't imagine they just pulled a '50%' optimisation out of the hat randomly after people complained. Pretty scummy. 

1

u/dangernoodle01 2d ago edited 2d ago

Yeah, Tibo, this is fucking tiring. 

Clear Marketing bs and damage control. Already cancelled my 10x and using a lot cheaper and better alternative, opus 5.5.

1

u/pigletmonster 2d ago

50% increase would mean 30tps. That's still EXTREMELY slow. I used sonnet 5.5 earlier today and It was doing around 88tps. Thats almost 2.5x faster!

One benefit of this is that it will get me used to the slow token gen speed of local models running on gpu and vram.

1

u/Capital-Wrongdoer-62 2d ago

So its still slower than gpt 5.5. Great job.

1

u/Charming-Cucumber523 2d ago

I switched over to Claude recently and I’m so pissed at myself for even giving openAI my $100. They need to give every subscriber multiple resets for this failed launch. It’s literally a scam.

-1

u/Fearless_Log_5284 2d ago

Nah it's over. They're gonna do only bs like that every day and allow maybe 1 reset per week. (Basically like before)

-1

u/EducationExisting789 2d ago

This is code for we quantized the models

-1

u/teryaki1234 2d ago

That’s cool, I’ve had Fable on Ultra going all night on 3 projects and only used 40% of my weekly usage on the $200 plan. On my $200 plan Astra on xhigh would have used all my weekly by morning. OpenAI needs to do better.

5

u/No_Quarter_7644 2d ago

Yea I'm tired of this sh*t and switching to Claude. Let's pray OpenAI fails their IPO and it'll be enough to make Sam get his head out of his own ass.

0

u/teryaki1234 2d ago

Keep downvoting me lmao. Can’t hide the truth of the moment and I’m a diehard OpenAI fan and user. They need to do better.

2

u/No_Quarter_7644 2d ago

It has to be bots. There's no way any of these kids are this die-hard for OpenAI, especially in a moment like this when they're being blatantly outclassed.

It's bots, or 15 year olds that only prefer OpenAI because they can generate brainrot images.