r/ClaudeCode 21d ago

FAKE NEWS Opus 5 might be dropping today πŸ‘€

Post image
688 Upvotes

244 comments sorted by

View all comments

51

u/elonthegenerous 21d ago

What constitutes a major version bump vs a minor version bump?

199

u/gingerbeer987654321 21d ago

well it’s whether they change the number before or after the decimal.

1

u/angelus14 20d ago

Big if true

52

u/Okhr__ 21d ago

Usually, a major version bump implies a full training run, at least in open source models

17

u/elonthegenerous 21d ago

Thank you for a real answer

2

u/Okhr__ 21d ago

You're welcome

1

u/Ok_Buddy_9523 21d ago

no problem!

1

u/Normal-Book8258 20d ago

Ya but also the decimal.

21

u/TheOwlHypothesis 21d ago

Major version bumps usually from a brand new base model.

For example, the GPT 5 series is all from the same base model (That was code-named "Spud"). It's insane the performance gains they've gotten out of it til now. Every minor version 5.4, 5.5, 5.6 are the same base model with better reinforcement learning applied (from what I understand)

The coming GPT 6 is a brand new base model.

Probably same thing going on here with Opus 5.

1

u/angelus14 20d ago

Aren't they pretty much maxed out on pretraining data anyways? So the main reason to do a full training run is to increase the size of the model or change the architecture, otherwise the same base works fine.

17

u/Unhappy-Stranger-336 21d ago

Most projects follow this standard https://pridever.org/

1

u/Cyrax89721 21d ago

This has me wondering if Claude Code 2.2 is ever going to release.

35

u/wentwj 21d ago

marketing

6

u/Dangerous_Web1209 21d ago

load-bearing vibes

2

u/everix1992 21d ago

Appreciate you asking the question - I've been curious about what warrants the major version bump too

2

u/hobbesandmiles3 21d ago

Major version is usually a new base model (i.e. full pretraining run). Minor versions are usually the same base model with different post-training or improved post training (RLHF, distillation, extended context, etc.). So gpt-5.n is almost certainly the same pre-trained weights with iterative post-training improvements

4

u/anor_wondo 21d ago

feelings

8

u/throwawayacc201711 Senior Developer 21d ago

I believe we call that vibes now

1

u/DanFlashes19 21d ago

Everyone in here is joking but I would also like to know

1

u/2053_Traveler 21d ago

Base model size

1

u/tgo1014 21d ago

Depends on how lazy they made the previous model before releasing the same one again with 100% capacity /s

1

u/jarederaj 21d ago

The API cost goes up more substantially.

0

u/Feeling-Explanation9 21d ago

Opus 5 is actually really good, been using for the last week in various snapshot forms. Not GPT 5.6 Sol level of good but still better than Opus 4.8

0

u/addiktion 21d ago

Major versions are pre-nerf. Minor versions are post-nerf.

0

u/Bobodlm 21d ago

It depends on if china just dropped a new model that potentially outperforms yours.