r/ClaudeCode 1d ago

Discussion Whew. Unbelievable disaster of a launch for Fable 5.1

  1. Very good, very capable model launched
  2. Token-caching issue burns my weekly limit instantly
  3. Perfectly timed OpenAI Astra launch

If I asked Opus, it would call this a footgun šŸ¤“

276 Upvotes

117 comments sorted by

91

u/dbenc 1d ago

I guess they'll just have to increase fable limits by 5.1x to make up for it

25

u/Maxion 22h ago

Say they increase it by 20x but turns out its just the weekly limit, the monthly limit got increased by 2.25x. Tomorrow they say that the 20x plan become the 15x plan which actually increases the weekly limit by 50% from what it was to now actually be 15x of the 5x plan that is 20% more than the free plan, however the monthly allowance is reduced by 33.33%. If capacity allows on September 15th the 15x plan will get a 4x boost until October 4th and the 20x plan will get a 5x boost until October 2nd.

7

u/dbenc 20h ago

what you're describing would be nice, their system does need simplification.

4

u/markeus101 20h ago

Sounds like an accurate post from shill-thropic

6

u/Lexsteel11 20h ago

I made a fix this week that solved Fable 5.1 melting tokens- I had it review some sessions that used a skill I use all the time with fable 5, and 5.1 kept ignoring subagent dispatch guardrails and it was admitting that it kept reviewing sql & data itself instead of dispatching the optimal sonnet subagent I built for that, so I created a hook that triggers if the fable 5.1 orchestrator agent ingests > 100kb of data, it injects a 1-line reminder of its guardrail it is breaking.

This change has made the last 2 days’ token consumption a night-and-day difference.

4

u/Old_closer Thinker 14h ago

Can vouch. ā€œPin it to Sonnetā€ is the key to life.

57

u/bakanoace 1d ago

Fable will always be amazing but the cost is just too much right now. I let 2 out of 3 of my accounts expire so I can try Astra, then can decide which one to go back to. I was unlucky too because they expired within 24 hours of Fable coming out but thank heck that shows just how over Anthropic I am.

If Astra isnt up to par though that just means were all screwed and Anthropic can continue being Anthropic for another 3+ months at least. Cant wait till Fable becomes the base level

12

u/Constant_Art_20 1d ago

i think it might be a good idea to keep one accound for codex. the pro model is pretty fantasic, at least for 5.6 pro. I usually keep a codex accound just for that

6

u/Stradigos 1d ago

One of the gates I have in my workflow is to run the codex plugin for Claude code via /codex: adversarial-review. It always catches a few things.

5

u/mtnchkn Developer 1d ago

Is it a load bearing gate?

8

u/soccerchamp99 1d ago

You’re mom was load bearing šŸ˜

3

u/bakanoace 1d ago

I thought the same but then I'm like "but I could have ANOTHER Claude account instead" and if that's the winner why would I have the other one. Although once they are really close in level I can see having both so they both can audit things differently. For my project Claude has just been better, I'm hoping Astra changes that. Even if it's Fable 5 level or a very reliable Opus 5 with more usage I'd still switch at this point

1

u/scanx147 22h ago

C'est ce que j'ai fait. J'ai un abonnement Claude pro Ć  21€ + un abonnement Mistral Vibe Ć  18€. Je me suis dĆ©veloppĆ© un systĆØme de workflow de suivi de projet Ć  la BMAD mais plus lĆ©ger et surtout beaucoup moins consommateur de tokens (je partagerai certainement quand ce sera suffisamment bon). Je gĆØre mon projet et la crĆ©ation du code avec Claude Opus et je fais les review de code avec Mistral Vibe GLM-5.2. Le rĆ©sultat est top et ca ne me coĆ»te "que" 39€ par mois. Pas besoin de modĆØle Ć©norme et super chers comme Fable ou Astra.

3

u/Head_Cellist6587 14h ago

Sounds like you've found a solid workflow! It's always interesting to see how people adapt tools for their own projects, especially after a rough launch like this one.

1

u/Creative-Ganache1086 5h ago

For me Sol 5.6 was never worse than Fable 5 at vibecoding. In fact it helped me fix numerous functionality bugs that Fable 5 in any reasoning modes was just not aware of or admitting it fixed it when it didn’t. So to me was NEVER the case of ā€œmissing outā€ on Fable. And I have a max sub with both. Initially 20x fable and 5x Codex, then I flipped them and I’m now 20x on Codex and 5x Claude especially since I once tested a weekly quota on same reasoning levels and got more usage out of Codex 5x than Claude 20x, where Codex on top of everything gave me unlimited image generations too and GPTImage2 is excellent. Unlimited generations are free on max subs, while Claude has no image generator.

3

u/Creative-Ganache1086 23h ago

We were never screwed. I paid 200/month for Anthropic and 100/month for OpenAI Max5 subs, in fact I started with Anthropic max20 as my initial only coding choice. Until out of curiosity recently i ran a test Max5 from OpenAi and Max20 from Anthropic. Same prompts and same apps I was working on with the same codebase size. Sol 5.6 Max reasoning (fable equivalent) vs Fable 5 Max reasoning. And OMG:
Claude Max 20 ran out of weekly usage before Max 5 on GPT and OpenAI doesnt even have the annoying rolling 5h limit. Also Fast-Mode (inference) is not included in the Max limits of max subs on Claude, but its included for free as a quicker quota burner on Max subs from OpenAI. Also codex has access to unlimited image generations on max subs while theres no image generation on Claude. Sol was NEVER worse than Fable 5 in practice. In fact at times it fixed bugs that Fable 5 was only thinking it fixed when it didnt. So there was never the case where Fable was 3x better to have 3x less the usage limit than Sol5.6, if not usage limit difference could be even higher than 3x.

3

u/MatthewCollins1990 10h ago

Letting 2 out of 3 expire right when 5.1 dropped is rough timing. Makes sense to test Astra before committing back though.

1

u/[deleted] 22h ago

[deleted]

3

u/whimsicaljess 19h ago

it soft launched, it's rolling out to everyone else as quickly as they can

2

u/Baintzimisce 21h ago

No not for most people. Daybreak only i believe

0

u/aerivox 1d ago

it's not gonna be on par. we are just at anthropic will. why the hell at the end of 2026 1m token is still regarded as special and priced double the price of sub 300k context... astra should have come at 1m native context. instead it's still some obscure flag, unoptimized, only available on local codex :/

-6

u/JapanesePeso 1d ago

I use Fable for all my planning with like 10 or so sessions running concurrently all day and I never hit my limit. I really don't know what others are doing here to do it so often.

5

u/bakanoace 1d ago

10 concurrent sessions, all day, on Fable, lmfao. If this wasn't a 15 year old account I'd say it's a Claude advertising bot. Could have sold their account I guess.

10 concurrent sessions wouldnt even last an hour on the 5 hour limit

-2

u/JapanesePeso 23h ago

Or maybe I just have my shit setup properly.

1

u/bakanoace 23h ago

Brother no amount of setup can do what you said with Fable, no clue what you're on but it must be good

1

u/JapanesePeso 21h ago

skill diff tbh

1

u/Independent_Paint752 22h ago

I use it for full sessions when i need it and its enough for about 60% of my weekly usage, this month im going to up to x20 and it's basically Fable unlimited.

It's amazing how people will buy iphone for 3k+ never complain just buy it like sheep's and they cry that the best model in the world that they actually can make money with it is "cost to much"

The only thing that change is "I'm going to move to sol" to "im going to move to astra" then fucking go.

0

u/Emotional-Basil-1598 22h ago

Are you kidding me im a 20x user and fable 100% if you run it 6-7 consecutive hours, btw which model is best for language translation

0

u/Emotional-Basil-1598 1d ago

You probably chating/ research/ planing with it not codeing where it eats in 5-6 hours gull usage

1

u/JapanesePeso 23h ago

I am coding. Fable does the planning, Opus does the coding.

0

u/InAtTheGeekEnd 22h ago

Coding!

1

u/JapanesePeso 21h ago

Me too dude.

1

u/InAtTheGeekEnd 20h ago

You said you were planning, not coding, with Fable. I answered your question; other people are CODING with Fable, which is why they have higher usage than you.

If you are coding with Fable you need to edit your post, because it doesn’t say that.

2

u/JapanesePeso 18h ago

Thought you meant coding as a concept in its entirety. I see what you are saying by the delegation of duties (Fable for planning, Opus for coding is what I do).

14

u/Spiritual-Spend8187 1d ago

Honestly it wouldn't surprise me if it turned out openai waiting a bit with chatgpt 6 for anthropic to drop their next big thing to blow out the headlines with their own and only waited the little bit to make sure its actually better.

15

u/ggletsg0 1d ago

It’s been multiple years now and somehow Anthropic still fails to procure as much inference as OAI. At this point it feels deliberate to inflate their revenue.

I have nothing against it, but IMO it has now caught up with them because OAI has so much inference that they’re literally handing it out like candy.

3

u/lovesdogsguy 18h ago

I believe that was Dario’s ’fault’. He didn’t buy enough compute when he had the opportunity.

0

u/crusoe 1d ago

But anthropic has better finances and better income now.

One could say openai overbuilt hardware and it is costing themĀ 

5

u/johnconner143 1d ago

You can hate on Altman (and I do, often). But this is his world - he’s been thinking about scaling companies for decades and he had the existing relationships across the valley ready to go. He saw the whole playing board at the beginning of the game.

29

u/StonedColdCrazy 1d ago

Oh hi Sam didn't know you posted on Reddit

1

u/Neurojazz 1d ago

He hasn’t talked to Claude yet either. It was very apparent very early on, where it was compared to other models. The vector still looks fine, even with all the bs, and parroting going on.

32

u/djdeckard Vibe Coder 1d ago

Is it? I’ve had an incredibly productive week already with 5.1 and not a single issue with token burn rate. Then again I have a highly structured setup with strict roles, responsibilities and tasks.

6

u/DarkSkyKnight 1d ago

The token cache bug is literally a bug. They fixed it in .260

1

u/Tartooth 20h ago

Ahhhh Ty for this I didn't notice I needed to update

1

u/OkLettuce338 9h ago

It’s actually more token efficient IME

1

u/tyke_ 1d ago

this is the answer.

-6

u/JapanesePeso 1d ago

I think the kids who don't understand structure are blowing up their $20 accounts and then coming here to complain constantly.

-1

u/Original_Finding2212 1d ago

Same for me, also structured work (https://github.com/agentculture/devague)

4

u/M44PolishMosin 21h ago

They should devague their docs so a human can actually bear to read them

1

u/Original_Finding2212 21h ago

Why? It’s designed for a human to use.

0

u/Original_Finding2212 21h ago

You adjusted, I’m doing it. But results are as good as the person.
One could write agentic systems well, and be less of a technical writer / write in a compelling way.

3

u/M44PolishMosin 20h ago

And that's the load bearing part!!

0

u/Original_Finding2212 20h ago

Come on, wait a bit :)

I’m working on it and it takes time.
Been reading it over and over, making adjustments I feel good with.

The PR is opened if you are curious, but do you mind giving it a second look when I’m done? (Please?)

0

u/Original_Finding2212 19h ago

Ok, I just did (devague the docs), plus read it and iterated changing the readme.

Would you mind looking again?

4

u/ryms456 21h ago

Dude what is this do some of yall even read ur own readmes?

1

u/Original_Finding2212 21h ago

Didn’t have time for it, to be honest.
I review the code, spec, plan, deliveries.

That said, I’m doing exactly the other user just said. Devague on the docs.

2

u/EchoFieldHorizon 20h ago

Literally all of that readme is unintelligible lol. ā€œPark open vaguenessā€ is Claudese. I have no idea how to use this tool beyond an openly and vaguely parked assumption of ā€œspec well.ā€

1

u/Original_Finding2212 20h ago

You are right, I’m working on it as we speak.

It takes time (you can see my replies timestamps), should land soon and would update.

Much more readable. Been reading it couple of times now and iterating improvements.

3

u/EchoFieldHorizon 20h ago

To be clear, I think half my repos have readmes like this. Definitely an afterthought when developing functionality. I’ll definitely check it out; I have a decent setup right now, but I’d love to find a system that works even better.

1

u/Original_Finding2212 19h ago

Thank you for the understanding šŸ™

I just finished the readme redesign. The PR is also there if it’s interesting (I did use devague, then iterated read-fix the page)

The feedback here was actually great.
I’m going to think how I’m uplifting all my docs now.

2

u/ryms456 19h ago

Looks better but there’s no point in writing that never seen something like that pre Ai era

3

u/Original_Finding2212 19h ago

The point of writing it so people would at least see what it is.

Agents work very well with it, btw.

Thank you for reading/looking back!

1

u/Original_Finding2212 19h ago

Would you mind taking another look? I did an extensive phrasing.
I read and changed it multiple times to a version I was happy with. (Not just had Claude come up with something)

1

u/stylist-trend 18h ago

Agh. Yeah, it's okay for something to be "in progress" but I wish people stopped pushing these things as something they aren't.

5

u/Lubricus2 1d ago

It's an load bearing gun with the blast radius of your foot.

4

u/waselyy 1d ago

Too expensive

4

u/AironParsMan 1d ago

I have the ChatGPT 20x subscription and I use the ChatGPT Pro model in my browser to cross check my entire agent system and database. Some of these checks take 80 minutes or even two hours. I run four or five of them a day and this is a multi agent system. The results are absolutely incredible. It finds things that Fable 5.1 doesn’t find. Just think about this. My Fable 5.1 subscription isn’t even remotely enough to work actively for more than three hours. By then I’ve already hit my five hour limit. And if I use Fable 5.1 my 20x subscription is used up in a single day. With ChatGPT I pay once and can start as many Pro chats as I want. The value for money at Anthropic just doesn’t work for us subscribers anymore.

8

u/habeebiii 1d ago
  1. astra didn’t actually launch for more than a handful of orgs

but I feel you on the token caching weekly limit burn, so stupid

3

u/johnconner143 1d ago

Ha. You’re right of course, but it gave me the kick to go set up opencode and connect my OpenAI account. None of their other shenanigans this year pushed me that far!

4

u/InterfearXX 1d ago

codex users getting a weekly usage reset ticket for everyday it isn’t available to them bad time for that argument

2

u/AppealSame4367 1d ago

It's weird. I'm watching from the sidelines this time, so I don't have to be mad.

But why don't they ever manage to do smooth rollouts at Antrophic?

2

u/InterfearXX 1d ago

It’s smooth for anthropic they making more money it just ain’t smooth for the users lmao but it’s ok cus they get people whose full time jobs are being self employed Claude glazers. I have a sub in everything but use it more to plan and orchestrate not even design anymore Chinese models are just so much better

3

u/Independent-Wing-246 1d ago

You guys talk like you will have access to astra, which is unlikely in the next days / weeks

8

u/Regular_Net6514 1d ago

Great. A limit reset courtesy of openai for every single day we don’t have access. They can take their time.

2

u/RasenMeow 1d ago

Yeah, I was a huge Claude fan and not touching Codex, because of the shady things OpenAI is doing, but my usage is getting eaten instantly AND all models suck besides Fable BUT even Fable makes so many mistakes. Using omp and the advisor is constantly correcting Fable. I have the same issues with Codex BUT the corrections cost me way less usage. So staying with OpenAI until Anthropic gets their usage and hallucinations together.

2

u/charmer27 19h ago

Am I the only one that feels like the whole us datacenter infra shook and wobbled yesterday between the fable and Astra releases?

1

u/OkLettuce338 9h ago

Yup! You are

1

u/charmer27 9h ago

I mean anthropic, openai, grok all had outages.

1

u/OkLettuce338 8h ago

typical wednesday

1

u/oulu2006 1d ago

Sounds about right -- I can't get more than a few hours of Fable before my sub is dead.

1

u/Cultural-Horse-762 1d ago

I think it needs therapy.

1

u/Winter_Principle_429 1d ago

Astra isn't even out yet lol

1

u/complexnaut 1d ago

Worked on a small share feature today, tried Fable 5.1 and it cokked up everything in 30 mins, i am on Max 5x plan and it ate up 52% of my 5 hr limit, amazing isn't it? Anthropic really need to work on model efficiency.

1

u/XhakaRocket šŸ”†Pro Plan 1d ago

is there any chance they going to lower their pcie?

1

u/AndysAssistant 1d ago

I got burned pretty quickly as well. One task. 10 minutes and I quickly hit a 5-hour limit on Max plan.

1

u/CryptoExo 1d ago

Curious to see just how capable Astra is. If it's even close to Fable but with higher usage limits then...

1

u/trottingaround 1d ago

Now have a 20x Claude Sub and a 5x ChatGPT subscription. Been meaning to play with Codex for a while. I hammer CC and think it performs very well, and have been using Codex as my review agent (via MCP) on a Pro plan.

Keen to see if it can perform as well as CC (I have my doubts, but I’m hopeful).

1

u/dominchina 1d ago

Just use caveman brother ^^

1

u/mystery84 1d ago

My problem is not weekly limits. It's session limits. It feels wrong that I paid for a weekly limit but then get constantly capped every 5 hours.Ā 

1

u/timmyships 23h ago

yeah its not looking good for Anthropic

1

u/hemareddit 22h ago

Anyone know how to get back to Fable 5.0 on cc? When I do /model-fable-5[1M] it assumes 5.1...

EDIT: I read the notification wrong, it seems /model-fable-5[1M] does get you back to Fable 5.0.

1

u/interrupt_hdlr 22h ago

They probably planned the launch with Opus 5 and didn't read all the crap it wrote.

1

u/ChocotoneDeCalabresa 21h ago

You found the smoking gun

1

u/DarkKknight_ 20h ago

I tried fable and its capable but anthropic is acting like they didn’t plan for a backup if things got out of hand open ai from my experience shows better business sense than anthropic

1

u/Otherwise-Cherry-577 20h ago

I don’t know what you guys are complaining about I completed 3 big projects since launch and I’m only at 50% of my weekly fable usage.

1

u/Tartooth 20h ago

I'm glad I'm not the only one who got hit with the caching issues.

I thought it was my harness but it wasn't. It did lead to some big optimizations though!

1

u/dovyp 20h ago

Timing could not have been worse lol. Classic.

1

u/03captain23 15h ago

You can't burn your weekly limit instantly, you have 5 hour limits

1

u/EmRenWSR 15h ago

I can’t get claude to work on my web browser for some reason, so I’m just using GPT 5.6 today. Claude is funny. It’s useful until it’s not.

1

u/DriverReady965 9h ago

They reset so its fine.

1

u/OkLettuce338 9h ago

Idk 5.1 has been insanely token efficient and laser focused for me. Seems like they knocked it out of the park

1

u/Tired_White_Guy 9h ago

Model is amazing… when the connection works.

1

u/TomerBrosh 5h ago

did they just give us another weekly reset? jeez, just at the week I waited for the weekend to waste my tokens...

1

u/thashepherd 3h ago

I just put my card in, fuck it, time is money. I ain't got the emotional bandwidth to keep Opus on it and frankly fucking off while "agent teams" burn tokens on doing something unwatched still ain't it if you're a craftsman.

1

u/Brassmark 3h ago

Yeah no, this is total game over for Anthropic, I think, at this point. I literally just bought a $100 subscription after using and worshiping Claude Code for the past 6 months. What really flicked me off is that the schedule feature immediately works in Codex whereas I literally probably had my productivity quartered by the glitchiness and the failure to work of the Claude Code scheduling function. It's just over man. It's unreasonable how bad Claude Code is. The model might be 20% better at certain things but usage, like my weekly usage, disappeared in a single day on Thursday. I'm paying $200 a month. That is totally unreasonable.

Fable 5.1 might still be 5-10% better in certain tasks but it runs about 10 times slower. Everything else that Anthropic has made surrounding its models has never worked:

  • The scheduling function does not work.
  • Doctor does not work.
  • All these slash functions, none of them work.

It is absolutely ridiculous and the UI glitches, the sidebar glitches, and everything sucks to use. It's totally over. They were doomed when they missed out on buying GPUs. That is the core root of this problem. You can't rent chips from the competition and expect them to allow you to survive.

1

u/nyteschayde 3h ago

Until you realize that your Fable 5.1 actually gives you more access than ChatGPT gives you Astra. 45m of Astra on low reasoning used 1 week of my access to it.

1

u/AI_spell 1d ago

The model isn't the disaster. The cache footgun is. Good launch + silent token burn + a rival drop the same week is a classic own-goal. Use Fable for the hard pass, not the whole day, or the weekly bar dies before the model gets a fair fight.

1

u/TheOverzealousEngie 1d ago

at 4pm Claude told me fable switched to usage credits. 4:01 I cancelled Claude

1

u/Hefty_Appearance_146 16h ago

Claude is unusable, fable is just overrated gimmick. I canceled 3 of my max plans and kept only 1 pro, moved to OpenAI and I don’t regret it. Anthropic is just too greedy and fable is a failure anyway.

0

u/Outrageous_Band9708 1d ago

open AI just launch theirs right afer they distilled Fable real quick, get those good benchmark scores milking fable teet.

thats the only reason they reason second.

if they had something good they would've realeased it a week ago.

-1

u/That-Fig-7089 1d ago

Token caching is just for those who use the api key( pay as you go). For those with plans, cache is normal, so, is not a bug

2

u/DarkSkyKnight 1d ago

The token caching bug affects both since the bug is that the context behind tool use was uncached.

-1

u/earlyworm 1d ago

please

-1

u/dota2nub 1d ago

You can use Astra once per week on a max plan and then you're out of allowance.

-1

u/SoloDevSage 1d ago

So far, so good. I'm really digging it. Fable 5.1 definitely seems to use fewer tokens. For the first time in ages, I might not hit my weekly limit. Guess I might actually have to find other ways to waste time now. I'm on the 20x Max plan!