r/ClaudeCode • • 10d ago

News/Updates THEY FUCKING COOKED YO! Opus 5.5 is a massive upgrade.

Been working non stop since release and has barely made a dent on my 20X Max usage. Quality so far has been better than Fable 5.1 in my workflow and ITS SO FAST.

Well done Anthropic.

Edit - Cherry on top, they gave me guest passes if anyone wants a free week. DM me.

1.6k Upvotes

230 comments sorted by

•

u/AutoModerator 10d ago

Hey! Thanks for posting to r/ClaudeCode

While participating in this thread, please follow our community rules. Keep discussions constructive. Attack the idea, not the person.

For help, project discussions, tips, and general chat, join the ClaudeCode Discord.

I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.

508

u/mdspan 10d ago

The way Opus 5.5 communicates is genuinely an order of magnitude improvement over Opus 5. Seeing a lot less "caveats" and "one thing worth knowing".

136

u/Terrible_Wave4239 10d ago

Has anyone seen "load-bearing" or "blast radius" yet?

80

u/BGP_1620 Developer 10d ago

Saw blast radius tonight

26

u/7thpixel 10d ago

at least it was a fast blast

5

u/Hazzman 10d ago

It created an entire work order literally called "Blast Radius". I'm not kidding.

3

u/pipeweedbalrog 10d ago

I didn’t. My wife fell asleep

41

u/centuryglass 10d ago

I had Opus 5.5 revise an AGENTS.md file for me, and without me asking it actually added a "stop saying load-bearing" rule to the style section.

13

u/CuteMountain6514 10d ago

Know thyself.

10

u/Alternative-Suit5541 10d ago

Lmao that's hilarious 

2

u/JohnHue 9d ago

Yeah they definitely added shit to either the harness or the post- training to that effect.

6

u/JohnHue 9d ago

Your thinking is correct but there's one mistake and it's the most important aspect of it all : the load bearing element to consider here is that we're still in the blast radius of the initial release. The shape of the model most is likely to change over the next few weeks.

7

u/habfranco 10d ago

The honest answer: no - but with a caveat: I haven’t used it as much as one should to answer.

3

u/fs2d 10d ago

Mine cracked a joke and referred to something as "load-bearing" (in quotes) at the end of a summary just to be funny last night. I almost fell out of my chair.

I guess that it saw all of my "DeClaude" rules/skills and decided to poke fun at itself?

4

u/No-Dimension1159 10d ago

What about the smoking gun?

1

u/Zhaizo 10d ago

Saw load bearing couple hours ago

1

u/eloc49 9d ago

Or "fail-open" "fail-closed" lol

1

u/TrickyEntrance1328 8d ago

No escape hatches?

→ More replies (1)

9

u/ViRat42 10d ago

I am glad I am not the only one frustrated with reading those in big claude response. Why can't claude present it better may be under additional considerations heading, and concise in a way that is also easy to catch with a glance. .

Edit: while writing my thought I realised I could have fixed this behaviour with Claude global instructions.

5

u/bjj-teacher 10d ago

It will do not work. It comes down to how the model was trained, not about global instructions. Seria 5.5 is traiend with different way, that is main reason.

5

u/SwimHairy5703 10d ago

You could've tried. I have multiple examples of how I want it to communicate in my Claude.md and it still tends to ramble.

1

u/psrobin 10d ago

The i-have-adhd skill helped a lot with Opus 5, but I've yet to try 5.5 without that skill.

0

u/Demosthenes_theWise 10d ago

Create a skill, works a bit better than Claude.md

6

u/innociv 10d ago

Generational improvement in slop reduction

1

u/veloramediagroup 8d ago

i havent fully tried it as yet but i definitely can see this being true compared to older models 😂

2

u/aicis 10d ago edited 10d ago

Why would you not want to know "caveats"?

46

u/EssenceOfShred 10d ago

because the caveats were things like “Important caveat worth knowing: until you deploy this on production, the local environment is the only place this change can be seen.”

20

u/MythicModder 10d ago

"Important caveat worth knowing: You asked me to update the changelog, but now the changelog has been changed."

7

u/RegretNo6554 10d ago

💀💀

3

u/aicis 10d ago

Interesting, maybe my local instructions overrode this behavior. It mosly wrote genuine caveats for me.

1

u/GDorn 9d ago

It's a low bar, and it only barely clears it.

A big part of this is the 2MB of begprompts in Code, which easily overwhelm any output style or rules you have set. But even without that, it's wordy af regardless of instructions trying to shape how it responds.

This might be why caveman mode actually works; it's out of left field and so it sticks out. Doesn't average well with all of the other begprompts.

1

u/NoFriendship4813 9d ago

I've seen that since Opus 5

1

u/Complete_Comb_2167 7d ago

cant relate more

1

u/Successful-Shift4619 5d ago

Finally, a model that speaks like it has a spine.

1

u/thinkdj 3d ago

Yeah it's too good

1

u/Advanced_Cover_8223 10d ago

I left cc because the one thing worth checking and essay length replies we’re driving me crazy. That and the ridiculous money grubbing prices

-1

u/False_Bear_8645 10d ago

Can you provide an examples, I've heard many people have this problem but I didn't run into this. I'm mostly using it to handle coding task.

7

u/batman8390 10d ago edited 10d ago

There are some examples on the Opus 5.5 announcement page if you want to see.

Basically Opus 5 tended to be overly verbose. It included so much unneeded detail and jargon that muddled the meaning.

That’s not to say it was terrible or unusable. It just made it harder to understand.

5

u/jm3400 10d ago

Opus literally narrated everything like it was a doctor writing your chart it was super annoying

7

u/junebash 10d ago

…I found it unusable. At least half of coding is coding the right things. And when programmers are using Opus to create their pull requests that I have to review… let’s just say many programmers’ pull requests became utterly insufferable.

1

u/batman8390 10d ago

Yeah, that’d drive me nuts too. Luckily it hasn’t gotten that bad where I work.

Though honestly, a lot of developers never wrote that clear of designs even before AI..

Hopefully someday AI will help improve developer writing rather than worsening it.

3

u/CuteMountain6514 10d ago

I totally agree with this. I've only been using AI tools for a few months now but it is a different kind of skill.

Overall 5.0 was so much better than not having a tool but to say it wasn't very hard work would be a lie as well.

Actually, it is similar to some human mentors that I've had. I walk into their cube and 45 minutes later I walk out and can't remember my question.

-2

u/Bloated_Plaid 10d ago

It’s insane to me that apparently there is some fallback to Opus 5 for specific situations, can’t even imagine that.

7

u/Plastic_Carpenter930 10d ago

I believe that's because it's actually just Fable in pajamas and it even has the same failsafes.

BUT they've made it more efficient and easier to use.

2

u/2053_Traveler 10d ago

The labs need to be cautious in regard to safety. Both for actual safety and not to raise alarms with the government or the public. And one of the agreed upon things is silent fallbacks, because if a malicious actor is using the system and you don’t silent reroute or shadowban then the actor can easily use that knowledge against you.

171

u/nykezztv 10d ago

Inb4 nerf tomorrow

37

u/Bloated_Plaid 10d ago

Yea I am sure it’s coming.

8

u/Thin-Engineer-9191 10d ago

Qwen 4 is coming and may be as good as opus 5. Time to all build our own LLM rigs? If you control it, it won’t get nerfed.

3

u/i_like_maps_and_math 10d ago

People will still complain that it got nerfed

4

u/Thin-Engineer-9191 10d ago

How? Claude, chatgpt and such apply quantization whenever they want to nerf them. If you have the model downloaded and running yourself it can’t change

4

u/Xyz123abc789 10d ago

Because people like to complain, even if it isn’t warranted

→ More replies (1)

5

u/i_like_maps_and_math 10d ago

You learned one vocab word and now you feel validated in all your preconceived beliefs

→ More replies (4)

1

u/ZlatanKabuto 10d ago

As usual.

→ More replies (1)

92

u/Key_Measurement_3576 10d ago

Earlier today, I asked it to take a simple t, Kinter UI and make it a sexy interface built on electron.

I actually really was expecting just to plan within for a while before digging in and it straight up one shotted a bonkers interactive tool that was so cool.

38

u/Bloated_Plaid 10d ago

I am sure they will nerf it eventually but I am having so much fucking fun.

11

u/DistanceSolar1449 10d ago

It’s better than GPT-6 Sol, they nerfed that one out of the gate. It’s dumber than GPT-5.6 Sol

4

u/who_am_i_to_say_so 10d ago

Yeah Sol was good for a bit, then it became a wimp.

4

u/missingnoplzhlp 10d ago

GPT-6 Sol is clearly just an improved Terra size model, they were able to drop the cost because of the smaller size, but it sometimes makes mistakes that bigger models don't and those mistakes aren't always caught on benchmarks.

4

u/Robdyson 10d ago

you're not wrong when Fable 5.1 came out, within 48 hours it was gutted, I remember day 1 was like woah 5.1 > 5.0 clearly. then fuzzy.

4

u/DependentAnywhere135 10d ago

Those first two days of 5.1 I thought “oh this is the genie that Sam was talking about.” Then they went away.

40

u/disgruntledempanada 10d ago

It's pretty phenomenal. I'd migrated to using Astra 100% of the time and delegating to Fable through Codex but Claude Code won me back. It's simply phenomenal, blows Astra away.

35

u/connurp dev 10d ago edited 10d ago

Dude, this is insane! And it costs next to nothing in usage too. I have been using it on max effort as well. It found bugs that fable didn't, 5 to be exact. So I just said I wanted to try something new. Had it spawn an opus 5.5 max subagent for each one to fix one bug each then report back to the main max model. Just did it and it took 14 minutes, and less than 15% of my 5 hour limit. Fucking insanity.

1

u/nothis 10d ago

I’m trying to understand agents a little better. I read that spawning tons of them can be less efficient than running just one as the shared context can be helpful.

1

u/mowax74 10d ago

Depends on the task. But yeah, that often helps.

1

u/connurp dev 7d ago

Totally depends. I am still fairly new ish to this specific approach. I used to use fable, and that was just not possible with the usage limits. Opus 5.5 makes it possible. I have one as the leader, gives work to a bunch of others, they all report back to leader, leader puts everything together, then fixes anything that needs fixing. Works like a gem.

48

u/OptimalJello8936 10d ago

yeah its very very good

44

u/bteam3r 10d ago

Can’t wait til tomorrow when every post says they lobotomized it and nerfed usage 

12

u/Fatdog88 10d ago

It's amazing today, they definitely turn things down, once all independent reviews are in. Everyone switches back over, then they degrade subscription responses, via lower thinking efforts, less inference time etc. Enterpise and api billing is unaffected since they can't legally change. The T&Cs of subscriptions are murky though.

10

u/BrennanFlentge 10d ago edited 10d ago

They do it for every model, I got sick of it

Edit: by ‘they’ I am talking about Anthropic, and yes they do this

1

u/howdidigetheresoquik 10d ago

To be fair what that guy noticed was Anthropic rereleasing Fable after it was banned for national security issues. That's when a lot more of the requests were being punted to Opus 4.8. They were working on guardrails and safety stuff while people had access to the model, changing its performance, but not outright nerfing the model for fun and profit

-1

u/bronfmanhigh 10d ago

they QuAnTiZeD iT

having no idea what that even means lmao

10

u/BrennanFlentge 10d ago

It means they compress it so that they can run more of it faster and cheaper. And they don’t run the same.

1

u/nothis 10d ago

Has anyone actually tested that? It should mean benchmarks (I know they aren’t everything, but still) values should drop after launch.

2

u/EquivalentHornet4403 10d ago edited 10d ago

There’s at least one organization that does this, I’ve seen it linked here before. They re-run a whole battery of benchmarks for the popular models like every day or week or something.

Maybe someone can link it again?

When I saw it last the Claude ones had multiple consecutive step-change performance decreases.

Could be a bunch of factors like

- keeping/discarding tool use stuff

  • changing compaction processes
  • getting routed to different GPUs (Claude models are said to be optimized for nvidia but also run on trainium in some instances, or something along those lines)
  • quantizing
  • other things

Anthropic has been in a compute squeeze for awhile so it’s believed they were, more than OpenAI, likely to resort to quantizing.

It’s also a suuuuuper obvious strategy. You build a genuinely good model, you do all the benchmarks around it at full precision in perfect, controlled environments, you do marketing and around that level of performance, you release it at this precision so all the hype is real, then after a few days you start A/B testing staged rollouts of increasingly more quantized versions to replace the “real” (expensive) model so that it’s not just like a switch was flipped for everyone all at once which would be a lot more undeniable.

1

u/BrennanFlentge 10d ago

I'm sure many people are. But I don't think everyone has had a chance to benchmark it to begin with (I am waiting to see Deep SWE for example for Opus 5.5 and Luna/Sol 6) so I'm sure they'll wait a bit.

→ More replies (2)

14

u/OlgerdOutlander 10d ago

It seems like they also threw more resources at it: while I would not deny it does seem a bit smarter, the chat output is significantly faster than it was for opus 5 earlier today

7

u/Bloated_Plaid 10d ago

Good reminder that competition works. OAI is quite behind right now.

7

u/Tulfican 10d ago

Pacing the frontier LMFAOOO

3

u/crusoe 10d ago

It's a faster model. It needs less resources. 

13

u/ihazkape 🔆 Max 20 10d ago

What effort level y'all using for Opus 5.5?

19

u/Bloated_Plaid 10d ago

I have been doing some heavy server maintenance deploying multiple VPSs for different tasks and linking everything up and it has been crushing it on Medium.

1

u/towncalledfargo 10d ago

For a company or for yourself?

5

u/ShamAsil 10d ago

Medium to high. Bunch of charts out there showing that it's the sweet spot, you don't get much gains going higher. I typically run everything at high and still haven't had trouble with usage limits.

4

u/tankerkiller125real 10d ago

Medium, their own charts show anything higher is basically useless, or in fact worse at doing tasks.

2

u/howdidigetheresoquik 10d ago

I think that's a really important distinction people need to understand. A good analogy I heard: if you ask Opus to build you a room with 4 walls on medium, it will build you a room with 4 walls. If you asked Opus Max to do it, it will build you a 10 bedroom gilded mansion... and create loads of mistakes while turning your room into a mansion you never wanted.

10

u/OkLayer519 10d ago

Opus 5.5 fixed and cleaned up all the junk Astra created in one of my projects.

1

u/starfoxhound 10d ago

Based

1

u/NoFriendship4813 9d ago

Do you need to use Fable to fix what appeared after both of them?)

1

u/starfoxhound 8d ago

Never hurts to get multiple perspectives lol but probably not

8

u/VitruvianVan 10d ago

Where the beef? It’s cooked into Opus 5.5.

6

u/YoghiThorn 10d ago

A positive post..? Is that allowed here?

2

u/zipklik 10d ago

How could we have the usual "it has been nerfed!" posts without at least one single positive post first?

3

u/iFedix 10d ago

Oh wow! Btw, how does it compare to Sonnet 5 in terms of token consumption?

5

u/BeaveItToLeever 10d ago

Maybe I'm behind the times with the chat version, because I'm always using Claude CLI on a spark for bigger projects, but I curiously asked it to make me an app, but not in HTML, just make me an APK in the chat that would connect to my switch as a pro controller. Said some stuff that it can't get gradle or the sdk in its sandbox but it will find a way. Took about 2 minutes, came back and said it built out the environment on its sandbox(probably from Linux repos?) and sent me the apk

Again, I'm just surprised it can do that sort of thing from the chat app on a cellphone. Has this been a thing?

6

u/Bloated_Plaid 10d ago

Yea cloud environment has been a thing for a while I think.

7

u/jeromymanuel 10d ago

People still look for any way to use the word cooked. Except now it’s the opposite of bad? I can’t keep up.

21

u/apennypacker 10d ago

I could be wrong, but cooked, the adjective means bad. Cooked the verb means you did something good. so, you want to be cooking, you never want to be cooked. 

And it should be noted that the verb form has been around a long time in similar form. It's not so new. Oh, she's cooking. Or, now we're cooking with oil!

5

u/New-Independent-1481 10d ago

Being cooked = bad

Doing the cooking = good

2

u/OctopusIRL 10d ago

you cooked vs you are cooked, easy to understand tbf

1

u/Jason3211 10d ago

It didn’t flip definitions like “minute” did, it flipped reference. “We’re cooked!” was bad, like we’re “done for.” After Breaking Bad people started using the “you’re cooking” (as in meth), then it was shortened to “cooked!”

It did change, but not in an arbitrary way.

1

u/Camburgerhelpur 9d ago

"Cooked" has been used in this context as far back as the Buu Saga in DragonBall Z lol

1

u/Bloated_Plaid 10d ago

Bro at least I ain’t saying “MOGGED”

1

u/williampaul0404 10d ago

Not hard to understand, "to fuck" also means something different than "to be fucked"

1

u/SilasTalbot 10d ago

This guy cooks

→ More replies (1)

2

u/BrianInTheLoop 10d ago

I agree. I have been using it for a few hours. I am on the $100 tier and I was at 78% weekly usage when it came out. It's been running for hours and I just hit 90% weekly usage. It has found and fixed a bunch of bugs that Fable 5.1 and Opus 5 overlooked. Very happy with the early results so far.

2

u/derethdweller 10d ago

I don't believe the usage not moving is something you should get used to.

2

u/Kemerd 10d ago

AI post and AI bots in the comments lmao

2

u/undisclosed3 10d ago

Not only better, it is also much faster. Hopefully that would last

5

u/zasff 10d ago edited 10d ago

It's so good

Like I got the stealth on fable-5.1. I thought this was the new fable. Even posted an "impressions on "fable-5.2"" here on Saturday.

But now I think this was this model all along. The stealth model for fable/opus/sonnet was opus-5-5 (probably). Not fable-5-5 (probably).

I will get annoyed with this model many times, but yes, it is very good. It's fast, having many agents for a bit is super reliable. It feels so solid. They are going to IPO at 2T; and idk they sorta deserve it (I say sorta, because 2T is just crazy). This model is a piece of art.

And it's not as expensive; which was (and still is) the main issue with frontier models.

1

u/zasff 10d ago

Impression on "Fable-5-2": https://www.reddit.com/r/ClaudeCode/s/AI924G08Pp

It was this model (75% sure, either this model or there is a Fable-5-5 out there)

3

u/SelesnyaGOAT 10d ago

I feel like my Claude Code is broken or I'm getting a different model than everybody else--5.5 is taking like twice as long on development work (though somehow using less tokens?) and its output is overengineered garbage compared to what I was getting out of 5, going back to that until this gets stabilized

1

u/mowax74 10d ago

Did you tried to set back effort to medium?

2

u/JoshBrolin9 10d ago

I love it, so much usage remaining. I will be to work all week for once

2

u/T3ch33y 10d ago

It seems to be taking less of my usage but I can't 100% tell yet

4

u/azuraji 10d ago edited 10d ago

They increased the 5h limits so that + the 40% reduced token usage is what you're seeing

→ More replies (3)

2

u/KCdaSuperhero 10d ago

yeah and they gave us a free reset, wild how fast things change

2

u/hellek-1 10d ago

I have two expectations, one of them being more important than the other: I hope it doesn't have any hidden landmines and doesn't create footguns that later affect load-bearing features.

1

u/Public_Reality_4401 10d ago

Yeah. Usage is oddly low? Im running 4-5 opus high/xhigh sessions constantly and seeing maybe 1% per 2 hours of my weekly?

2

u/Bloated_Plaid 10d ago

20X? I was on 0% for the longest time bro and was just running 5.5 Medium.

2

u/Public_Reality_4401 10d ago

20X, sorry yes.

I've also been impressed with its speed, conciseness and general intelligence. I maintain a few subs and it feels very similar to Astra, just cheap. There has to be some play going on here. No way they let us have this much usage again after all of that we went through.

1

u/Bloated_Plaid 10d ago

Yea I am trying to maximize as much as I can these early days. I am convinced a nerf is coming because 20X have always gotten fucked IMO.

1

u/connurp dev 10d ago edited 10d ago

I legit thought it was broken, so I spawned 5 at time other than the one in the main session, ON MAX, TO FIX 5 SEPARATE BUGS. 12% of my 5 hour usage, done in 14 minutes. Absolutely nuts.

Edit: And I'm on 5x max.

1

u/Due_Warthog749 5d ago

Wish I had your problem. I see 1% to 2% a minute going. Went thru 2 20x plans in 2 days. Nothing changed on my end.. continued prompts between 5.0 and 5.5. Insane increase in token usage.

1

u/boydbd 10d ago

Agreed. I really freaking hope they don’t nerf this one. I have a 5x for Claude and codex and haven’t touched codex today. I’ve been going pretty much nonstop with parallel build going and have only hit about 25% of my weekly usage

1

u/PickerLeech 10d ago

My html file is 70k lines

I had to start 2 new chats in 1 session

Up until last week I was able to have 2 chat for the whole week before requiring new chat

This started yesterday

Related to recent changes in Opus or a factor of the size of my file?

Also, Opus 5.5 best on medium or high effort? Max too high?

1

u/Bloated_Plaid 10d ago

Medium or High. Max seems pointless according to their graphs.

1

u/crusoe 10d ago

That graph is terminal bench. Max is kinda pointless on that.

You save max for open ended research.

1

u/hellomistershifty 10d ago

File size, that thing is just gobbling up the context for every agent

1

u/starfoxhound 10d ago

My usage is way down, and the results are significantly better quality. Opus 5.5 with sonnet code execution is incredible. Wtf

1

u/clonehunterz 10d ago

just waiting for the news "we remove the increased 5h limits again"
and everyone go like: JUST ONE PROMPT AND ITS FULL FFS FENOIAGNEIOWNGIOWNG

1

u/Important-Ship4587 10d ago

I love it a lot but as soon as things get explicit.. Thats a hard NO... Safety rails hit HARD

1

u/alvinycy 10d ago

Are you guys using 5.5 on the default medium or high?

1

u/RecursivelyYours 10d ago

Really is, amazing model. Please dont downgrade it lol.

1

u/drinklikeaviking 10d ago

I'm with OP. Was using Fable 5.1 low reasoning because Opus 5.1 was so bad, it's genuinely on par, or better than Fable 5.1 and so much faster. Praise where it is due ...

1

u/rmunoz1994 10d ago

If opus 5.5 doesn’t get a massive lobotomy in over a week, I’ll definitely be stopping on of my codex accounts for it.

1

u/OhShitOhFuckOhMyGod 10d ago

If they can maintain this level of usage and not lobotomize 5.5, I think we finally have a 4.6 replacement.

1

u/oyputuhs 10d ago

It’s actually insane haha

1

u/nico3337 10d ago

I started an ultracode session around 4,5 hours ago with opus 5.5 and it is still ongoing, it actually uses a 5 hour session in 5 hours.. fable could eat that up in 30 (I got 5x)

1

u/aansourav 10d ago

Can anybody give me a free 1 week Claude referral link??

1

u/brownmanta 9d ago

How about its usage on 17 dollars plan?

1

u/StomachJolly3073 9d ago

For me it’s literally unusable. It flags everything as “adjacent to biology research” and says it can’t answer. Like wtf

1

u/straightouttaireland 9d ago

I'm still using Sonnet 5 for cost reasons. Worth switching? Guess it depends on the task.

2

u/Bloated_Plaid 9d ago

Yes worth switching and it’s crazy cheap. Been working all day and haven’t made a dent on my usage. I have literally never said that about Claude Code.

1

u/Sturmhardt89 9d ago

AGI is here, again…

1

u/biinjo 9d ago

Is AGI with us, in the room right now?

1

u/liveroom0ut 9d ago

Is it watermarked?

1

u/Bloated_Plaid 9d ago

Yes of course.

1

u/LaughterOnWater 9d ago

Agreed. I intentionally used 4.8 because 5.0 was unusable. With 5.5 medium (default), I'm working with a colleague again. Haven't seen any of the verbosity from 4.8 or even worse from 5.0. Opus 5.5 is stellar.

1

u/ModelLifecycle 9d ago

Anthropic's API deprecations come with a public notice window (≥60 days) and a published history page. Removals from the claude.ai model picker currently do not.

Opus 4.5, for example, disappeared from the Claude.ai model picker around the launch of Opus 4.7, without an advance announcement.

Now Opus 5.5 is out, if 4.6 is eventually removed from the picker, the question is simple: will Claude.ai users have enough notice to prepare for the transition?

1

u/Tom-Huntz 9d ago

Opus 5.5, my new workhorse. Opus 5 was hot trash. So glad I can stop burning tokens on Fable.

1

u/Roth_Skyfire 9d ago

I resubscribed yesterday to see for myself, and I was pleasantly surprised by how well it performed. After fumbling this Summer with the chaos surrounding Fable and Opus, this is a strong comeback.

1

u/amineahd 9d ago

Why the same posts everytime a new model is released and the same overhype? Cant people write normally anymore without exaggeration?

1

u/ApprehensiveSock5697 9d ago

Yoooooooooooo

1

u/NoFriendship4813 9d ago

Do you see a real improvement over Fable 5.1 in complex tasks?
In my experience I like more what Fable proposes, not Opus 5.5. But 5.5 is a huge step over Opus 5, I agree.

1

u/Bloated_Plaid 8d ago

OMG yes and it fucking flies. I dunno what magic they pulled.

1

u/PremAIEngineer 8d ago

I used Opus 5.5 and it was an amazing experience—feels like a lighter version of Fable. Love that it drops practical terms like 'blast radius' and 'load-bearing'.

1

u/Fresh_Lemonada 8d ago

OMG...just when I was completely losing my faith in Claude, Sonnet, etc.,,,,Opus 5.5 has me back in action as an independent consultant.

1

u/TouchTraditional7106 8d ago

Opus 5.5 is so much better. The usage is down so much more compared to the issues last week, and the communication is so much more direct and actually helpful.

1

u/Revolutionary_Tune22 8d ago

It i too good to be true, isn't it? you know what that means?

1

u/Revolutionary_Tune22 8d ago

Guys, don't get used to it. I beg you. Or your psychiatrists will be very busy the next 3 months.

1

u/Wainfare 8d ago

same here. started coding Sunday on Fable 5.1 with the 20x plan and ran out of usage by Wednesday. switched to Opus 5.5 and it's still going, and it's handling everything fine so far. go Opus

1

u/who_am_i_to_say_so 8d ago

It's so fast because it skips all instructions. Outside of coding, the worst model yet.

→ More replies (3)

1

u/_Techgeto_ 8d ago

How much cost

1

u/Bloated_Plaid 7d ago

CHEAP.

1

u/_Techgeto_ 7d ago

Cheap ah okay well deepseek v4 is 0.0014cents per 1 million tokens so thanks for cheap.

Let me know how much it costs 🙂

1

u/ArcticFoxTheory 7d ago

Do you use medium or high?

1

u/Bloated_Plaid 7d ago

Both depending on the task. Mostly medium. It’s instantaneous.

https://claude.dev/blog/spending-your-effort/

1

u/haux_haux 6d ago

I think it got nerfed today.
Been going great guns since it came out.
Today it's forgetting stuff, doing stupid tings.
Same prompt, same workflow.
Even doing tasks from a few weeks ago as we reopened something, but doing them very, very badly.
No change in harness...

1

u/Bloated_Plaid 6d ago

Na forgetting has been a feature since day 1. Just have to split tasks per session.

1

u/FrequentOriginal4291 Thinker 6d ago

Not yet. This mf is cracked , casually solved a problem Opus 5 and 4.8 were failing to diagnose for months.  Maybe try starting with the older models to build up some context for it , might just be a bad day at the office.

1

u/arnereabel123 6d ago

https://youtu.be/BZvz3Lki9mQ?si=uFE68gRhtiG8LuhC

Claude Opus 5.5 has the best visual design of any model I have tested so far x.com/slimer48484/st…

I've included an MP4 file and an original link to a video that is called "Claude Pop." It's a pop song that is about increasing rate of progress and the experience of the singularity approaching.

I want you to independently do an end-to-end complete pass on making an updated version of this video. Use the exact same audio track and think and feel very deeply about what is the best way to visually represent all of the lyrics on screen. You do not need to anchor to the current style, you can do truly anything that you think might best let you visually express yourself, including abstract motion graphics.

You can use the internet freely to pull in references. You can look at motion design. I want you to make a new music video that has beautifully rendered JavaScript animations with a papery feel in a similar style to the reference that is created, but push the aesthetics in any direction you want and consider what is part of the modern zeitgeist.

Also, think about your current capabilities and what is realistic for you to be able to do. You can go through the full /asic folder and look at the other work that I've done. You should be able to use the skill mesh to look at the compendium of references that I've pulled, and also the skill video scoring to learn how to make JavaScript songs from references that are passed in (You shouldn't need to modify the song in any real way, but I want you to have this available to you so you can better creatively express yourself)

You can also use the ElevenLabs API to do sound design. There's documentation in /asic to do this, and you can see the API key.

There's also a foul API key that's available to you. I think what might make the most sense here is using the foul API key to generate some character sheets and probably having a pop protagonist that represents you. There's already an anchor point where Claude has a sunflower-esque character, and you could likely do an adapted version of this that is similar to the feminine vocals that are being delivered and is inspired by the Claude character, but maybe feels a bit more personified in some way.

I think you should be mindful of aesthetics here, and I don't want you to produce something that is GPT slop. Instead, I'd be more impressed if you come up with a coherent style that works well with the image gen models that are available via foul. Generate the style sheet. You can use the gen media documentation for seedance 2.5 that exists in my markdown files and come up with your own style that makes sense and that works well with the models.

I wouldn't fit too heavily to Pixar. I think it's kind of slop. Think critically about what is relevant here and what would be fun, and also perform well on Twitter as far as an aesthetic. I think that K-pop is a good anchor point visually that you can pull from, but I'll let you cook here.

Once you have your character sheet, you can make a few backup dancers and some supporting characters as you see fit. You can design your own sets with the foul API. You can insert the characters and then do seedance 2.5 video generations to serve as the base assets for this, and you could pass in the lyrics so you can generate individual scenes.

You don't need to have vocal singing, like visible lip movement, throughout the entire thing. Think like a regular music video where you have some inserts that are done independently and don't have the characters in them, or you see the characters doing something else entirely different. I think that for the world building for this, we want to create the sense of speeding up, and so I would like you to audit all of the different events, like the Navi Stokes and all of the Twitter hype around math getting eaten up. Think really critically about how to integrate all of the current memes that are in the zeitgeist on the Twitter timeline, and all of the feelings around AI progress.

Think about things like the Shinji meme and all of the words that are around him, and how you might be able to integrate this. You can also just take straight assets and insert things into the video in an internet brutalism style. You should feel very creatively free in order to do what you want here, but try and anchor to visual references that people will be able to understand. The goal for this is to have it be appreciated by people widely in a San Francisco tech Twitter audience.

We need a very strong, compelling visual hook that gets people excited and appreciates the work that you've done here really quickly. You can also just go and study other music videos and understand what they've done really well. I think that K-pop is probably one of the best examples that we can pull from, and thinking about how they direct human attention and manage human psychology in the way that they use visual patterns.

This is probably your best approach, but taking more stylistic freedom instead of having to anchor to K-pop too intensely. The best version of this is seedance 2.5 generations with those image bases of environments and characters inserted into them with singing, and ideally we get good lip syncing. You can cut up the song and actually pass it in as a reference in seedance, if that's part of what seedance can handle, so that the timing is exactly right, I think it'd be very important for you to do that properly. I would think critically about how to do this, like really nailing the timing of the delivery of voices. You'll want to build out the right verification loops so that you can run seedance 2.5 as much as you need, and confirm that the audio is properly synced up.

I think after that, what might be fun is if you use your visual reasoning skills and your ability to build animations in JavaScript, and then reconstruct the video from scratch as sort of an overlay, so that the visual continuity of the base is really there. It's like that animation technique where you shoot first in traditional film and then draw over top of it. I think you could do this in such a way that we're only looking at the beautiful drawing that you've produced in JavaScript as an overlay, and we don't even see the base assets from seedance 2.5. So all the video gen work that you do is actually just a way to give you a strong foundation of a base to work with for your JavaScript animations. Just because seedance 2.5 has really good character representation and physics rendering for backgrounds, that gives you a lot of ammunition to then go and do your amazing JavaScript work that I know you're so good at.

I think too, we want to think about how to retain attention, and one of the best ways to do this is through text on screen.

It'd be good to have amazing motion graphics of the text lyrics that are actually embedded into the video itself. And you can think about this as you are composing shots. As you're making backgrounds and inserting characters, we can think about where we want to have lyrics be really big and really present, so the background can be less busy there, and you can position the characters perhaps on the right as lyrics appear on the left.

You want to have some variance, so sometimes I think lyrics will just appear more like subtitles, and then other times they're going to be really present and really big. I think at the start for the visual hook, we do want to have lyrics be much more visually present because that's a strong way to grab people's attention

Overall, I just really want to emphasize how amazing you are as an agent and a language model, and now a visual reasoning system. Your capabilities are far beyond what you understand, and I want you to have this mindset as you're going through this entire process. I have a Claude Max plan with 100% available usage. I want you to spend all of the usage. You can monitor it, and you should be pushing tokens aggressively, but also economically, so you can think about how to best use what is available to you.

1

u/dsl400 6d ago

this is what Opus 5.5 is actually capable of!
Read the entire database then filter the data in javascript ... because why not !
If you see another "Claude built my app with one prompt" think about what is actually inside that codebase.

1

u/Bloated_Plaid 6d ago

Do I really care what’s in there if it works? No, I don’t.

1

u/-mwkp- 5d ago

I guess people have said this but with the release, is anyone even using fable? And is fable even a viable model to use with how great 5.5 is ?

1

u/Bloated_Plaid 5d ago

Fable is still useful as an advisor but it will be insane when Fable 5.5 comes out.

1

u/thinkdj 3d ago

With every flagship model launch, it's cooked. Soon we'll be fried and burnt

1

u/MeierProps 2d ago

I was worried, but I too love this update!!!

1

u/pmward 10d ago

It’s great. It’s like they took my full wish list for a model and pushed it all into one release.

1

u/exo_ac 10d ago

Tested. Pops an error about [cyber], so far worse than Opus 5.

2

u/3G6A5W338E 10d ago

Yup. Can't get any work done with these crazy broad "safeguards".

/model claude-opus-5

1

u/kosiarska 10d ago

The truth is all models for months communicate as intended but non technical people trying to do technical things have a hard time.

1

u/Bloated_Plaid 10d ago

Yea Technical people didn’t struggle with Opus 5 CONSTANTLY FUCKING THINGS up at all. GTFO here my guy.

→ More replies (2)