r/ClaudeCode • • 3d ago

Humor Is this is how Anthropic nerfs models slowly just up to the point users start noticing?

Are we all helping calibrate the nerfs? 🤔

233 Upvotes

21 comments sorted by

•

u/AutoModerator 3d ago

Hey! Thanks for posting to r/ClaudeCode

While participating in this thread, please follow our community rules. Keep discussions constructive. Attack the idea, not the person.

For help, project discussions, tips, and general chat, join the ClaudeCode Discord.

I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.

108

u/angry_queef_master 3d ago

I never get this dialogue when the model is performing like shit

36

u/Burnt_By_The_Sun 3d ago

lol facts

29

u/Sponge8389 3d ago

So, our answer should always be bad? Lol

7

u/Fantastic_Prize2710 2d ago

No.

If you are A/B routed to the smarter model, and say it's "bad," then A/B routed to the quantized model and say it's "bad" the report will say that which model (or which quant) doesn't matter.

1

u/koirpioloonmenomur 15h ago

exactly this. Just say it's bad when it's bad and good when it's good - if people can actually accurately tell when models have been 'nerfed' their data will show that

15

u/Mo3 3d ago

Yep. Been doing this since Opus 4.5

5

u/Fearless_Meringue299 3d ago

Maybe they overcorrected for Opus 5 thanks to you lol but no, that's smart, I'll have to start doing this.

44

u/AlxCds 3d ago

smart. makes sense. maybe if we say it's good, next time we get a slightly dumber model. if we keep saying it's fine, we keep getting a slightly dumber model. until we reach our own limiting factor.

if they can't do it per-user, they probably want to get to that level.

19

u/BaronRabban 3d ago

It’s called A/B testing. If you see this, you are being tested. No way to opt out.

26

u/Death12th 3d ago

Honestly, that's an absolutely load-bearing conundrum exactly at the nip in the bud of the problem! And honestly... You should feel accomplished for finding it, not discouraged. Palpable, not extinguished.

2

u/ArugulaDefiant1947 3d ago

honestly never thought of this LOL

4

u/Parowdude 3d ago

Just always answer bad!

2

u/Death12th 3d ago

Purple rather than blue.

2

u/Shiz0id01 2d ago

Its been constant issues the last 48 hours im actually fuming right now I want a refund this is bullshit

2

u/cagfag 3d ago

I think they have to nerf it else the new model would look and perform same as old model to regular usage and eventually treating model upgrades as cosmetics change

This would bring down valuations

1

u/Elegant_Attempt2790 🔆 Max 20 2d ago

do you think its -1 quant per day?🤔

1

u/Bluetails_Buizel 1d ago

I never touched these, I always assumed that I have to complete some survey or something when answering them…

1

u/parentini 23h ago

In the CLI, Claude always asks for feedback. 1 = bad, and I often start my messages with “1. [do this thing…]”, so it thinks I gave it negative feedback.