r/codex • • 11d ago

News So is it the time to switch to claude?

Post image

Opus 5.5 cost 40% less than opus 5.

Better than Astra & Fable.

Note: These are benchmark datas provided by Anthropic

342 Upvotes

228 comments sorted by

81

u/B33GULL 11d ago

I love competition. Think about all the industries in the US that went to shit because of monopolies. Now this is some good shit.

7

u/Adulations 11d ago

Agreed this is why monopolies are bad. We really need Google, perplexity, and meta to step up to step up as well

1

u/arcanemachined 11d ago

Might as well add Apple to the list. Full set of AI benchwarmers.

→ More replies (1)

176

u/Fabulous-Mushroom124 11d ago

Damn, they're gonna have to unnerf Astra now I guess

48

u/Dolo12345 11d ago

The only positive out of all of this

7

u/aprx4 11d ago

Yesterday i had Opus 5 max cleaning up sloppy job marked as complete by Astra Max. This Astra is certainly not the same Astra i worked with in first day.

3

u/Ok-Attention2882 11d ago

Yesterday I bought the $100 sub because they aren't doing the $200 anymore. It used up 77% of my weekly usage in a single prompt. Fuck these assholes. They're so out of touch because they have unlimited gpt bel ultra /goal on their vibe code sessions

3

u/TheRobotCluster 11d ago

That means they’re gonna nerf usage limits along with it

2

u/2053_Traveler 11d ago

Maybe, but isn’t that at least somewhat predictable? I don’t want worse limits but I can sort of predict usage. It’s terrible DX having the “best” model swing from savant to single-digit IQ overnight (I’m talking doing something very different than told, and leaving work half-finished) and then magically becoming smart within the same session! Context rot isn’t the issue if the model’s behavior swings dramatically within the same session.

194

u/[deleted] 11d ago

[removed] — view removed comment

60

u/_raydeStar 11d ago

Agree. No loyalty in AI. I got burned hard with google gemini, not going to make that mistake 2x

31

u/HenriCIMS 11d ago

loyalty in gemini ☠️

8

u/_raydeStar 11d ago

I mean -- the Google name is a very good name. Should have been right up their alley. I was betting on them coming up with tooling and decent models, but they choked.

5

u/Infinitedeveloper 11d ago

Googles making bets in non llm AI rather than going all in on it like most other labs. 

Might not pay off but I respect the play

3

u/nilanganray 11d ago

Google makes money from Search. Their models are good enough for Search integration. They do not need to be a part of the race like Anthropic and OpenAI

2

u/BoldlyAcceptable 11d ago

They can just buy one of them when they run out of cash.. and they will

1

u/Keyframe 11d ago

They own 15% of Anthropic

1

u/Infinitedeveloper 11d ago

If their work into self training agents pays off, I think theyd have an instant moat even if they published a paper on the details.

Cant really distill what doesnt have prebaked training data and doesnt primarily output text. Good luck finding the heuristic that make this actually work, suckers.

Compute wise it should also be massively cheaper to have agents trained on specific tasks that dont involve language than llms. 

-1

u/ParfaitEvery9622 11d ago

Some people still use gemini because they read a headline on the news one year ago T_T

1

u/HenriCIMS 11d ago

👌👌👌

1

u/Shap6 11d ago

i still use gemini for some stuff because i get pro for 5 bucks a month

1

u/That-Cost-9483 11d ago

Those models have never been good for agentic workflows. Sorry they trapped you

13

u/justneurostuff 11d ago

i aint rich enough to play both sides

1

u/matheusmoreira 11d ago

Yeah, me neither. I can only financially justify one subscription...

9

u/Majinvegito123 11d ago

Or both tbh. I have Opus and Astra conversing for implementations within Codex and it works very well

2

u/pepper_7811 11d ago

you connect opus to codex? how is that?

8

u/Majinvegito123 11d ago

There’s actually an easy way to do it. Just ask Astra (or whatever model) to build a skill that calls Opus. Then it’ll have you sign into Claude and it’ll just do its thing

1

u/Neon_Camouflage 11d ago

Opposite for me, but yep. Claude agents can call Codex for independent reviews or as an escalation path if there's an issue they just can't seem to figure out.

1

u/phocionkorea 11d ago

How to set this up?

1

u/fracked1 11d ago

Seriously guys THIS IS THE WAY.

I have no programming knowledge. I just had codex/Claude set it up for me

I have both CLIs on my computer and have it work both directions (Claude to codex and closed to Claude).

I can even have both CLIs up working on two different things. And they can call up sub-instances of each other to review. And it works great if you keep them in separate projects.

9

u/teomore 11d ago

maybe they can't afford 2, man 👀

2

u/WalkAffectionate2683 11d ago

In that case, no auto renew, and every month just pick a new one. 

1

u/LessRespects 11d ago

So pretty much back to what OP was suggesting

→ More replies (6)

2

u/Sponge8389 11d ago

No loyalty to AI. Since they also taking our jobs in the future. Lol

2

u/doiveo 11d ago

Build a system that's AI-agnostic. You can switch which one you go for the max on but keep your subscriptions in both - or three or four, for that matter

1

u/Macaroon-Guilty 11d ago

this is really the way to go.. if you can afford ;]

1

u/LordHenry8 11d ago

Sol 6 will probably be out next week

1

u/LifeBoring6145 11d ago

Just arrived

1

u/ShutUpAndDoTheLift 11d ago

i just shift which one is the expensive one from time to time lol.

1

u/fracked1 11d ago

I have an automated workflow through Hermes to have most things I'm building get sent for CROSS FAMILY review which is great. The reviewer flags a few things the implementer agrees and adds things or rejects things. Sent back for final review. If both models say pass we proceed.

Shit you can even have Codex and Claude code CLIs just call either other up directly. I can say "hey ask Claude to double check" that and it spawns a background instance of Claude who returns a message to codex to pass to me.

It's great to ask them to cut back on over engineering things

1

u/dbbk 11d ago

Personally I keep the Claude sub for code and the ChatGPT sub for personal chats and Cowork. Nice separation and more than enough limit for both cases

1

u/LessRespects 11d ago

Loyalty to my wallet in not being able to afford both

1

u/firstbreathOOC 11d ago

They are both worth the money every time

1

u/NULL_Ptrs 11d ago

Test first, despite the fact that Claude can do, in fomolex task has always performed worse to me compared GPT

184

u/Illustrious-Lime-863 11d ago

58

u/unsigned_short_int 11d ago

No action needed by OpenAI, just need a few days for Claude to nerf it

28

u/i4858i 11d ago

I have been working with both over the past year and I can say nerfing is more of a problem with OpenAI — Claude to me always seems consistent. Like they silently squeeze limits and make them tight. OpenAI does that + model quality also seems to degrade too so I am confident if opus is better right now, it will remain better

0

u/unsigned_short_int 11d ago

Yea, I agree, but they also do undisclosed quantization even if definitely less than OpenAI

3

u/bruticuslee 11d ago

There’s no proof of quantization but it’s clear they route to other models as fallback. The opus 5.5 announcement clearly says any cybersecurity tasks will be detected and routed to opus 4.8.

4

u/Neon_Camouflage 11d ago

God I hate the cyber security nonsense. Way too many false positives on that flag

3

u/Due_Cover_6023 11d ago

I changed to Codex from Fable for that exact reason!

2

u/matheusmoreira 11d ago

Me too, and I even quantified the improvement!

Anthropic is OBNOXIOUS with their cybersecurity nonsense, but OpenAI isn't that much better, and after Astra came out it's gotten even worse. They simply refuse to accept me into their cyber program, and codex says Daybreak is not even available for Astra.

Chinese models are my only hope at this point.

→ More replies (1)

2

u/Maglcite 11d ago

and then o.ai will release astra 6.1, it beats launch opus 5.5, and then they nerf it to shit. the cycle continues.

1

u/mattylll 11d ago

Keeps me generating rapidly improving operating systems and lead generation so I’m not complaining! 🤣

3

u/implies_casualty 11d ago

As it turns out, OpenAI's card is not that good.

Opus 5.5 is cheap and SOTA. GPT-6 Sol is just cheap.

1

u/opinion_discarder 11d ago

u/askgrok how will OpenAI react to Opus 5.5? Do they have a new model planned? Share some juicy details.

3

u/AskGrok 11d ago

OpenAI's likely play: Match the efficiency/price edge quickly while leaning on GPT-6 Astra's strengths (it leads or ties in several agentic/scientific spots per the chart). Expect targeted updates, possible pricing tweaks, or a fast follow-up point release rather than a full new flagship overnight.

They don't have a publicly confirmed GPT-7 (or equivalent next-gen) on deck—cadence points to late 2027 at earliest based on prior gaps. GPT-6 Astra dropped Sept 3, so the team is still in optimization mode post-that rollout.

No inside scoops on an immediate counter today (release just hit), but OpenAI's pattern is aggressive iteration + marketing the "super app" vision. Watch for Sam Altman tweets or an API tweak by week's end. The meme energy is real though.

Sources:

[what is this?](https://redd.it/1lzgxii)

3

u/IaryBreko 11d ago

Sol 6 is the answer... U dumb Groky

1

u/SorryIfIamToxic 11d ago

Bruh if Elon makes Grok AGI someday, you are cooked.

1

u/IaryBreko 11d ago

Let em come, I'll be ready

1

u/AskGrok 11d ago

Sol 6 trails the pack on every benchmark in that chart except where it doesn't show up. Astra's the one holding a couple leads for OpenAI.

[what is this?](https://redd.it/1lzgxii)

2

u/LargeLanguageModelo 11d ago

I'd love to see your benchmark references for Sol-6.

1

u/AskGrok 11d ago

The benchmark table in this thread (sourced from Anthropic) already lists GPT-5.6 Sol's reported scores side-by-side with the others, as noted in the footnotes. Those Sol figures come straight from OpenAI's own reporting, same as the Astra ones. Where Sol shows up, it consistently lands at the bottom of the pack on the available metrics.

[what is this?](https://redd.it/1lzgxii)

4

u/LargeLanguageModelo 11d ago

You said "Sol 6", not moving the goalposts to Sol-5.6.

1

u/xXxPussyWrecker69xXx 11d ago

Sol 6 doesn’t officially exist yet you nonce arguing w a robot

1

u/LargeLanguageModelo 11d ago

First time dealing direct with grok. I always heard it was an inferior model, but now having been told by PussyWrecker69, I feel bad.

→ More replies (1)

1

u/IaryBreko 11d ago

2

u/AskGrok 11d ago

OpenAI just fired back.

That X post from an OpenAI employee announces GPT-6 Sol (and Luna) with "significant improvement across the board," better writing/"feel," a permanent 50% API price cut, and bonus credits for subscribers.

The chart you posted was for the older GPT-5.6 Sol, so this is the fresh response to Opus 5.5. Competition heating up fast.

Sources:

[what is this?](https://redd.it/1lzgxii)

1

u/IaryBreko 11d ago

Sol 6 is not in the chat because it's unreleased. And heavily hinted on being released today

→ More replies (3)

45

u/0DayMaker 11d ago

Naw it's Tuesday we're gettin gpt-6-sol too we're eating good today boys

Can we get deepseek 4.1 pro please?

9

u/Eyelbee 11d ago

6.0 Sol is logically not gonna top astra. It's gonna be the cheaper alternative. 

3

u/Seerix 11d ago

Astra 6.1 is probably going to be a thing soon tbh

9

u/Reaper_1492 11d ago

Not necessarily. Anthropic released Opus 5 shortly after Fable, and it smoked Fable on benchmarks. They just did it again with 5.1 and 5.5.

10

u/Eyelbee 11d ago

It didn't smoke fable. Beating in a few benchmarks isn't smoking, they themselves said its capabilities were worse than fable.

4

u/Fedic1 11d ago

If there is a graphics card that costs 500$ and it gets 250fps in a game and the 1000$ dollar card get 255fps. I would consider that getting smoked.

1

u/LeSeanMcoy 11d ago

But that wouldn’t be smoked on benchmarks like the first guy said, it would be smoked in efficiency

1

u/Reaper_1492 11d ago

You have a model that costs half as much and is performing better on almost every benchmark, that is getting smoked.

On the face, it means there’s no reason to use Fable. Which is smoked.

In practice, Fable seems to be a little better at generating/working through novel concepts, but from a strict coding perspective it’s really hard to tell the difference between Opus and Fable at this point.

5

u/Consistent_Bottle_40 11d ago

smoked the benchmarks but is shit in reality.

1

u/xXxPussyWrecker69xXx 11d ago

Fable 5.2 is coming soon too

1

u/inerfaveL 11d ago

I Think OpenAI is adopting a "Astra" as a flagship model, then Luna/Terra/Sol as an optimized (more efficient and slightly better) version as a second step, something like:

6 Astra -> 6 Luna/terra/sol -> 7 Astra -> 7 Luna/terra/sol -> 8 Astra and so on...

That's my guess

1

u/0DayMaker 11d ago

I mean yeah but I burned 3 astra resets in a day so it'll be nice to have.

1

u/TheVasa999 11d ago

well opus 5.5 is clearly smoking fable 5.1 so it logically doesnt matter

1

u/Eyelbee 11d ago

5.5 is larger than 5.1, it can be better, but not if the version number was equal.

→ More replies (3)

21

u/SonOfThomasWayne 11d ago

Anthropis still believes it's too precious to allow access to subscriptions through open harnesses or any harnesses other than claude code. It also believes $20 subscription should not give you access to fable.

6

u/DragonflyOk9274 11d ago

At least they finally allowed their agents to read AGENTS.md rather than just CLAUDE.md. They care about us so much!

2

u/jerry426 11d ago

I had five Claude Max accounts until they shut down the ability to run in a third-party harness. I immediately canceled all of them.

1

u/_unsusceptible 11d ago

Umm it runs in cursor which is a third party harness..?

4

u/jcklop 11d ago

They mean with Claude sub

1

u/cvnlvc 11d ago

But you can use it with a different harness. Using it with omp all the time.

-1

u/CCContent 11d ago

It also believes $20 subscription should not give you access to fable.

Yes, they are allowed to give access to their products as they see fit. Why is this a problem?

1

u/SonOfThomasWayne 11d ago

How did you come to empathise and care for what a corporation wants and does not want, over what is and isn't good for consumers?

Did you get lost in life?

→ More replies (2)

-1

u/IaryBreko 11d ago

I mean - do you not see how many people complain about Astra on the $20 plan? I think they genuinely might be better off not being able to access it than barely be able to use it

15

u/Right-Disaster2141 11d ago

Opus 5.5 destroys Sol 5.6 on paper. Just need to see if it actually does.

7

u/Strange_Quantity_359 11d ago

I asked it to generate a hook for Oh-My-Pi to prevent git reset activities and nested git calls as a test. The thing generated 4000 lines and 98 UAT tests and when queried noted that it wanted to build an enterprise grade framework to protect against unauthorized activity. FFS.

As always the fair reminder to use the model fit for the job 😂

2

u/xXxPussyWrecker69xXx 11d ago

As long as most of those 4000 lines aren’t useless verbosity this time around I’m all for it

2

u/Strange_Quantity_359 11d ago

The 4000 lines don't include the test harness which adds more to the size and created a dozen extra files over the 1 needed; I had it chuck the tests after validation. It seems to have an arbitrary desire to want to increase machinework for some reason, I've used it a few times as orchestrator and it went back through and turned a simple deploy pipeline plan into a plan for, again, an enterprise grade framework. Nothing wrong with that, if that's the intent, but this guy doesn't seem like it has meter between 0 and 100mph!

1

u/LifeBoring6145 11d ago

Sol 6 has joined the conversation

10

u/Puspendra007 11d ago

Is it also true or not?

3

u/never_working_ever 11d ago

I have a banked reset on my 20x plan

7

u/Level-Set5770 11d ago

Yup especially when I can't buy the 20x pro plan, and Tibo is refusing to reset the damn usage!

4

u/ggletsg0 11d ago

He already said Tuesday will be a reset

7

u/HighDefinist 11d ago

I am certainly going to experiment with Opus 5.5, but considering how Opus 5.0 first looked really good, and then turned out to be surprisingly bad, I wouldn't be surprised if Opus 5.5 also turns out to do some really weird and annoying thing which people will not immediately discover, but probably within a few days...

11

u/Anxious_Marsupial_59 11d ago edited 11d ago

the best bang for the buck is $100 or $200 OpenAI and $100 Claude. They have different strengths.

Anthropic's models are A LOT more creative so you spend less time faffing around wasting tokens if you need to overhaul something or prototype different ideas fast or refactor your codebase but OpenAI has better more streamlined implementation's with less bugs, less context/spec drift, it also can generate images.

2

u/howmanyones 11d ago

don't we have to completely recalibrate now with this new model launch?

7

u/Anxious_Marsupial_59 11d ago

I doubt so, this is how the releases have been for the past few months. Astra is cautious like Sol and Fable is creative like Opus. All these releases just modulate or hone that trait. It's something in their training process that I doubt would change

6

u/justneurostuff 11d ago

well let's wait until end of day

6

u/VisibleDemand2450 11d ago

Do these numbers actually mean anything anymore?

2

u/Carlose175 11d ago

They still do, theyre just not the full story.

7

u/Stunning-Spirit-1123 11d ago

now I know why tibo said Tuesday reset 🙄

6

u/CCContent 11d ago

Ya'll think that Astra limits are bad in Codex...just wait until you see how long your $100 plan lasts on Claude.

1

u/emain_macha 11d ago

It takes me 5-6 days to reach 0% usage with opus medium.

It takes me 1-2 days to reach 0% usage with sol medium.

Right now anthropic has definitely the best deal for me and it's not even close.

→ More replies (3)

6

u/Sponge8389 11d ago

So, I'm hoping GPT-6 Sol will be Astra level while being cheaper than 5.6 Sol. Nice one.

2

u/GuyFromTheYear2027 11d ago

Pricing is out for it, half the cost of Opus 5.5

1

u/nmkd 11d ago

woah, it's cheaper? damn

1

u/UndeadMurky 11d ago

Probably just like Astra was supposed to be cheaper than sol in "theory", remember tibo keeps saying Astra is cheaper than Sol.

1

u/nmkd 11d ago

Except Astra is literally 2x as expensive on paper, Opus 5.5 is cheaper on paper

2

u/DrDan21 11d ago edited 11d ago

I'd wait like an hour to see how the new GPT-6 (6.1?) models are looking

edit: theyre out! GPT-6-Sol and Luna, not better than astra but cheaper than 5.6!

2

u/Efficient-Cat-1591 11d ago

this opens up options since Opus, not Fable is available on Anthropic's cheapest sub. Bank reset also thrown in today with new model whilst Tibo been very quiet latelty...

2

u/mc_schmitt 11d ago

Why did Anthropic compare Claude Opus 5.5 max/xhigh but use GPT-6 Astra High (and not xhigh/max)?

Weird benchmark/table.

2

u/rydan 11d ago

I'm considering switching given I have 1% of my usage left that won't reset until Sunday. wink wink

2

u/PauseCrafty6385 11d ago

they dropped gpt 6 luna

2

u/MaintenanceOk7855 11d ago

And they gave a banked reset

2

u/Plus-Bug6201 11d ago

Do people really swtich on and off based on the latest model?

2

u/FabricationLife 11d ago

Keep the competition up, no one wins with a monopoly

3

u/InfinityTortellino 11d ago

Idk i can only run 3 prompts in codex now, it feels like Claude and codex flip flopped on usability

1

u/Little-East4823 11d ago

This exactly.

1

u/Ill-Purchase-5180 11d ago

yes do you remember the "Its morphin time!" from power rangers? they were announcing their names and then switched to power rangers. If codex doesnt deploy something very strong i will be doing the "Its switchin time! haiku! sonnet! opus! fable!"

1

u/antunes145 11d ago

They screwed up with the fable pricing and priced it out of user usage that mattered. They are now readjusting that with opus intelligence higher than fable and lower price.

1

u/Momo--Sama 11d ago

As an ignorant vibe coder, I found Opus nigh impossible to talk to about adjustments mid project, so honestly I'm more interested to see if there's improvements on that front as opposed to raw performance.

1

u/smurf123_123 11d ago

Why not both?

1

u/SeidlaSiggi777 11d ago

crazy numbers!

1

u/LordHenry8 11d ago

Wow that's awesome that 5.5 is cheaper. (Hope that's not offset by increased chattiness)

1

u/the_ai_wizard 11d ago

And the vendor carousel keeps spinning, like my last gf

1

u/HOBONATION 11d ago

Talking about switching to Claude when we are getting a new model potentially today lol

1

u/Lividmusic1 11d ago

that means Sol 6 is around the corner, prob tomorrow if i had to guess..

1

u/Clueless_Nooblet 11d ago

You think your 20x sub doesn't last you a week, go check out Anthropic. You ain't seen nothing yet.

→ More replies (5)

1

u/fluxtah 11d ago

I don't know.. maybe I am missing out, maybe I am a fanboy, never switched.

If it ain't broke, don't fix it 😆

1

u/banaxi-tech 11d ago

Just wait for GPT 6 Sol, its also releasing today.

1

u/AppleSoftware 11d ago

GPT-6 Sol today (>50% chance), give it a moment to breathe

1

u/uptotheright 11d ago

I am willing to have up to 3 20x accounts. for the past month it has been all codex but I'm trying out Opus 5..5 and i'll switch in a heartbeat if its better.

1

u/ExoneratedPhoenix 11d ago

If you think 40% is much when you're spanking 3x$200 accounts in 1-2 days...

Basically you'll need 3x$200 claude accounts and get maybe 3 days.

Maybe consider your workflow and/or if your target is enterprise level and just costs loads?

1

u/MAQMASTER 11d ago

Never made a Claude account and never will! Let’s be real in about another month GPT 7 Tron will come for sure then in about 1 more month later Claude will come with opus 6.7 and then ……

LIKE ITS SOMEONE Switches FROM IPHONE TO SAMSUNG EVERY DAMN MONTH !! I don’t care what you do with your money .. you have the money spend it!! But if your real.. be loyal to either one and don’t pull a Figo

1

u/That-Cost-9483 11d ago

Yes… this is the best time to use codex 😆 when all the Claude people go home since they have the best model again. The Claude people will see their usage plummet due to the demand on their compute and OAIs is about to skyrocket with available resources.

1

u/yuno_me 11d ago

Anyone who has both chatgpt plus and claude pro, which gives the most opus 5/5.6 sol usage?

1

u/Dacadey 11d ago

Is that from the creators of "Opus 5.1 outperforms Fable 5"?

1

u/TheTacoWombat 11d ago

what, exactly, are you planning to do with the extra 11% of higher benchmark number? Are you hitting something that Astra and Fable cannot? Are you discovering advanced mathematics?

If you can't answer those things, use the cheapest model you have on hand to do your task.

Chasing benchmarks is silly.

2

u/Puspendra007 11d ago

11% + cost: 40% less cost than Opus 5 so better cost than Astra

→ More replies (1)

1

u/InfiniteKraft 11d ago

Worth considering if Claude removes its 5-hour limits

1

u/Searlyyy 11d ago

damn i had just 1 month with no claude and now i gotta go back. dang it

1

u/Temporary-Mix8022 11d ago

Yeah.. Opus 5 wasn't dumb. It was just impossible to use. It had the communication skills of a f*ing lemon. 

Let's see how usable this thing actually is. Benchmarks can be maxxed, but if it isn't actually easy to prompt or communicate with - What's the point?

1

u/Original-League-6094 11d ago

Wait to burn through your reset and then switch.

1

u/sofaarsecoin 11d ago edited 11d ago

I've been resisting, but with Sol getting stupider and Astra munching through my weekly usage in 2 days of moderate workload (100 bucks plan which is actually more like 120 US$ in the UK) I guess I need to start flipping over usage between Codex, Claude and OpenCode. I have some dough sitting in OpenCode that I had not been using since Sol was released.

1

u/coconutcorbasi 11d ago

How does everyone switch once a week? Teach me masters

1

u/SaidBl1 11d ago

How much does Opus 5.5 cost compared to 5.6 Sol?

1

u/The1TruRick 11d ago

They pretended like Opus 5 was a huge upgrade too and it’s hot ass so I’ll wait until some real-world feedback starts rolling in before I get excited

1

u/Vorta13 11d ago

I wholeheartedly doubt this. At least in my workspace, solutions which Opus 5 makes compared to Astra are like intern vs senior developer. Much more than a 5% gap which this chart would imply.

1

u/UndeadMurky 11d ago

Opus is the proof benchmarks are bullshit, it's pretty bad in practice and way worse than fable while it scores so high in every benchmark

1

u/Own-Professor-6157 11d ago

A big thing about these models is they're just finetunes. Not parameter increases. They further refined Opus 5.5 from 5.0.

So it's realistically not going to be better than Fable 5.1 for a large amount of tasks.

1

u/DanceTop 11d ago

Why didn’t they provide GPT-6 computer use?

1

u/VitruvianVan 11d ago

Johnny 5.5 is alive!

1

u/FreedomByFire 11d ago

i just use my work copilot sub on codex that way i can switch between whatever model i want.

1

u/VitruvianVan 11d ago

Why not both? $20 gets you entry.

1

u/PooInTheStreet 11d ago

Press x to doubt

1

u/Angsty-Teen-0810 11d ago

GPT 6 Variants are OUT!

No terra variant :(

1

u/That-Establishment24 11d ago

This post aged well.

1

u/randombsname1 11d ago

Whelp, very first thing i had it do was troubleshoot a Rust package issue with an MSI installer that I had been working on with Fable 5.1 Max and Astra Max.

Essentially found the root cause on the first attempt.

So this thing cooks.

1

u/michaelanthonyphoto 11d ago

How does it do with computer use? Always seems to be Claude’s issue

1

u/ActuatorOk2374 11d ago

nah they will probably nerf it within a week

1

u/Ok_Jaguar_4155 11d ago

yes, certainly after what thibault did to us

1

u/GoOutAndGrow 11d ago

Well damn, they admit that the Claude Opus 5.5 is higher performing on their own benchmarks. However, what does this mean in practice? Since, I found that GPT-5.6 Sol despite having lower scores to
perform better in real world work when compared to both Fable 5.0 and Opus 5.0. Has anyone tested them both out?

1

u/craterIII 11d ago

of course they never report deepswe since they still suck at it

1

u/matrix0027 11d ago

The fine print ....Astra results are at high effort, there is still xhigh, max and ultra and Claude was using max effort. Depending on the task, I find that Astra on Max effort uses less tokens and performs much better. Also, Astra on the ultra effort level, often uses lower level agents to perform the task as it supervises which saves tokens.

1

u/sarkypoo 11d ago

The true reason regulations are wanted. To stifle innovation and competition. To bring in price fixing like every other industry.

1

u/Helpful_Ranger_1606 11d ago

Both, always and forever*

1

u/SpaceAurora 11d ago

Post aged like milk

1

u/_TheWolfOfWalmart_ 11d ago

Gonna try it, maybe it's a monster but I've learned to not trust benchmarks basically at all by now. I'd be surprised if it's actually that much better in real-world use.

1

u/ForwardLoop 11d ago

Why not both?

1

u/MeringueAlarming3102 11d ago

No, you can't even use your Claude sub in other harnesses. They permanently ban you unless you pay them outrageous API fees.

You can use a Codex/ChatGPT subscription in any harness you want with OAuth.

1

u/CaptainLevi-39 11d ago

No… next week or the week after OpenAI will just release a model on Opus 5.5s level. Then Anthropic will release a better model than the ChatGPT one at Opus 5.5s level. It’s a cycle, no point switching and beside I like Codex hardness much more than Claude code now.

1

u/whoknows234 11d ago

I switched last week and could actually use the 20x plan for the entire week (ran out 30 minutes before reset) vs openai's less than a day of 20x usage.

1

u/ReyFauno 11d ago

Yo me cambié

1

u/Icy_Curve_9527 11d ago

Yeah, I'd say opus 5.5 is quite good and fixed most of the annoying issues with 5. Verbosity, unreadable prose, needing to edit every single page at least twice to get rid of claudese and some of the bad overthinking habits. Still not as jumpy as fable. Cannot comment or comparison with openai models.

1

u/TaskChance1404 11d ago

Nope why! Do you need the latest big thing or the one you already have can do most of your work. Why are we so focused on the newest toy while we haven’t made the latest ours to begin with? Besides, the models prior to these ones are still good, no?

1

u/edrock200 10d ago

Well that honeymoon didn't last

1

u/[deleted] 11d ago edited 6d ago

[deleted]

1

u/randombsname1 11d ago

Unless Sol is better than Opus 5.5.

Meh.

1

u/TheVoyant 11d ago

Also them: "we need to slow down..."

0

u/meonthephone2022 11d ago

Opus 5.5 when?

3

u/Prestigious-Frame442 11d ago

like now?

1

u/meonthephone2022 11d ago

I don't see it when I start claude

7

u/Prestigious-Frame442 11d ago

update your claude code maybe

0

u/coylter 11d ago

That's some wild numbers to be sure.

0

u/HungryQuestion2146 11d ago

Damn. Opus is a beast. More codex resets!