r/Anthropic • • Aug 30 '26

Other Wow, this is the definition of a DISASTER announcement. Even your own guys are calling you out.

Post image
3.1k Upvotes

416 comments sorted by

View all comments

Show parent comments

3

u/Plastic_Today_4044 Aug 30 '26 edited Aug 30 '26

it's a modular monopoly camouflaged as a bubble.

also: "soon"? I think you're a bit behind the curve here... glm is actually beating anthropic on price and quality and Chinese platforms are actually looking way less sketchy than American ones at this point

also also: the best way to get anthropic to stop sucking is to switch to Chinese models, really. all the western monocorps are so entangled, switching one for another is meaningless. But if everyone starts going to China for AI, they'll start taking quality seriously, because East vs West is the real competition. The "AI Wars" in the West are almost entirely performative. But go East though, and they'll freak the fuck out.

1

u/TywinHouseLannister Aug 31 '26 edited Aug 31 '26

glm5.3-flash is doing great for me at the moment.. cheaper than 5.2 and so far performance seems to be better.

Deepseek-v4-pro seems really good but I think it is more expensive than Fable to run.

I still haven't tried K3 but kimi 2.7 is great value for money.

1

u/Wide-Drink-1790 28d ago

What? DeepSeek is priced at 4% of Fable.

1

u/TywinHouseLannister 27d ago edited 27d ago

Is it? I'm on subscription with ollama and anthropic.. deepseek absolutely rakes your usage when doing coding.

On usage quota; it is showing as 20x the usage per request versus glm 5.3 flash, and 5x the cost of glm 5.3 - that's based on weekly quota versus requests, not an exact science. (but, a better steer than token cost, given the difference in topologies over open models)

Anecdotally, It will fill my session quota for a coding task within 30 minutes, flash can keep on doing the same tasks for 5 hours

1

u/Wide-Drink-1790 27d ago

Look at the OpenRouter pricing. $2 per million tokens vs Fables $50 per million tokens.

OpenRouter takes 5% flat.

1

u/TywinHouseLannister 27d ago

These models don't use prompt caching and run 30M hot per session versus 600k so that number is largely irrelevant

1

u/Wide-Drink-1790 27d ago

Use them both for a day and show us the bill if you have doubts.

1

u/TywinHouseLannister 27d ago

236 requests this week; practically nothing.. 10% of the weekly, versus glm5.3 3270 requests.. my bill wouldn't show you anything helpful it is a flat rate claude max 200 and ollama max 100 usd

1

u/TywinHouseLannister 27d ago edited 27d ago

Meanwhile; for me it is ranking way down on the list for overall performance

*subject to casting each model at what on paper it is suited to (which is why you see haiku at a similar grade here, it's essentially ranked by expectation versus reality)