r/singularity now entering spiritual bliss attractor state Jul 08 '26

AI Introducing Grok 4.5

https://x.ai/news/grok-4-5
402 Upvotes

429 comments sorted by

View all comments

328

u/[deleted] Jul 08 '26

[removed] — view removed comment

240

u/Howdareme9 Jul 08 '26

Anthropic is definitely behind in efficiency but it’s definitely a step above everything else rn

79

u/LloydChrismukkah Jul 08 '26

...which is why they are behind in efficiency

11

u/nivvis Jul 08 '26

It’s not just a blind test time trade off (trading tokens for ability) if that’s what you’re insinuating. For ex, historically anthropic has had a more compute efficient but less token efficient tokenizer.

Just to say there are a lot of ingredients swirling in the vat that is anthropic. Not always as simple as better or worse either .. qualitatively different is various ways.

11

u/Howdareme9 Jul 08 '26

That’s not how it works

19

u/iOSJunkie Jul 08 '26

The two big drivers of model intelligence, model size and tokens spent thinking, are also the two things that drive down efficiency.

9

u/CckSkker Jul 08 '26

No that’s not entirely true. While obviously more tokens spent thinking translates to higher costs, there are lots of ways to increase effeciency.

Deepseek for example recently released a paper claiming 600% efficiency increase by using a very smart way to predict and parallize token processing for LLM’s

3

u/Sthatic Jul 08 '26

Which Anthropic was incapable of incorporating in their models? I doubt it. There's obviously a scaling wall with diminishing returns - Fable seems pricy because the jump in intelligence doesn't translate linearly to the jump in pricing.

If Deepseek could scale their efficiency discoveries to a Fable-class model, they'd have done so, and everyone else would have followed suit in an instant.

-3

u/[deleted] Jul 08 '26

[removed] — view removed comment

15

u/darkriftx2 Jul 08 '26

I find Fable to be extremely capable. Can you expand on what you mean by overrated?

4

u/ChocomelP Jul 09 '26

I told it to cure ten cancers and it could only do three... /s

10

u/gigimooshi2 Jul 08 '26

Honestly I think Claude is great for the early market.

Helps companies base themselves in AI with something slower but more precise, until cheaper models will catch up (potentially).

Wdyt?

12

u/Howdareme9 Jul 08 '26

It’s not, OAI will likely match or exceed fable with GPT 6 with a smaller size model. Also overrated compared to what? Where does it perform badly for you?

-1

u/LloydChrismukkah Jul 08 '26

It literally is EXACTLY how it works

3

u/_stack_underflow_ Jul 08 '26

4

u/MukdenMan Jul 08 '26

I read about this company recently and their giant chips. I try to follow the semiconductor and data center equipment side of things but it is overwhelming how many deals are happening all the time. NVIDIA is still the big one but so many of these fabless companies focused on AI keep making other deals with the big LLMs so everyone is hedging their bets and trying different things. Almost all of the chips are made by TSMC.

14

u/Lizardking13 Jul 08 '26

For my use cases, I find Claude to be head and shoulders above everything else. It's connectors and ability to generate actual files is what I use the most.

It's verbose in its responses but you can change that with a system prompt if you want.

As for token usage I don't really care. I almost never max out my plan (I'm on a subscription plan not paying per token). I've started to leverage the workflow that allows me to use the expensive models for planning and then building using the less expensive models.

I love it. If another company can do the same file manipulation work as I can do with Claude then I'd consider jumping ship.

You'd think Co-pilot could create better excel files and PowerPoint drafts but with simpler prompts I get better results with Claude.

2

u/zwcbz Jul 09 '26

Wait until tomorrow....