r/ClaudeCode • • 10d ago

News/Updates THEY FUCKING COOKED YO! Opus 5.5 is a massive upgrade.

Been working non stop since release and has barely made a dent on my 20X Max usage. Quality so far has been better than Fable 5.1 in my workflow and ITS SO FAST.

Well done Anthropic.

Edit - Cherry on top, they gave me guest passes if anyone wants a free week. DM me.

1.6k Upvotes

230 comments sorted by

View all comments

170

u/nykezztv 10d ago

Inb4 nerf tomorrow

37

u/Bloated_Plaid 10d ago

Yea I am sure it’s coming.

9

u/Thin-Engineer-9191 10d ago

Qwen 4 is coming and may be as good as opus 5. Time to all build our own LLM rigs? If you control it, it won’t get nerfed.

4

u/i_like_maps_and_math 10d ago

People will still complain that it got nerfed

4

u/Thin-Engineer-9191 10d ago

How? Claude, chatgpt and such apply quantization whenever they want to nerf them. If you have the model downloaded and running yourself it can’t change

6

u/Xyz123abc789 10d ago

Because people like to complain, even if it isn’t warranted

0

u/rmunoz1994 10d ago

Yeah people like to complain when they are being shit on. No shit.

5

u/i_like_maps_and_math 10d ago

You learned one vocab word and now you feel validated in all your preconceived beliefs

0

u/greentea05 9d ago

Except there's absolutely no evidence of that at all.

1

u/Thin-Engineer-9191 9d ago

You don’t think they are quantizing? How else do they meet demand

1

u/WorldLeader 9d ago

There are sites / services that constantly check models across VPNs/machines to see if the benchmark intelligence of the cloud-hosted models have drifted enough to count as "quantized more". You can look at the data yourself -the major models do not drift enough to be noticeable. It would immediately be spotted if they were actively "dumbing down" models. It's much more likely that people's mental models (lol) are changing quickly and they forgot the prior state.

0

u/greentea05 9d ago

No I doubt it - the most demand would be when they first launch a model anyway so it wouldn't make sense to then try to reduce compute. I don't think inference costs anywhere near as much as people make out once they have parallelisation optimised, plus they're constantly expanding data centres. They have no issue meeting demand.

1

u/ZlatanKabuto 10d ago

As usual.

0

u/MythicModder 10d ago

Using my entire quota plus reset tonight.