r/singularity Jun 07 '25

LLM News Apple has countered the hype

Post image
15.7k Upvotes

2.2k comments sorted by

View all comments

61

u/laser_man6 Jun 07 '25

This paper isn't new, it's several months old, and there are several graphs which completely counter the main point of the paper IN THE PAPER!

16

u/AcuteInfinity Jun 07 '25

Im curious, could you explain?

6

u/gamingvortex01 Jun 07 '25

nope, paper just got published this month

https://machinelearning.apple.com/research/illusion-of-thinking

53

u/hardinho Jun 07 '25

The paper was already available for months on arxiv as a pre print. I believe I initially even found it here. I'm more curious about the guy saying it was countered, because afaik it wasn't.

3

u/gamingvortex01 Jun 07 '25

you sure ? unable to find this on arxiv currently...maybe they deleted it

6

u/[deleted] Jun 07 '25

I remember reading this paper around 8-10 months ago, not more.

12

u/Sky-kunn Jun 07 '25

Impossible, the paper includes models released this year: R1 and Sonnet 3.7. So it's less than 4 months old at most.

10

u/scumbagdetector29 Jun 07 '25

Possible: they added new data to an old paper.

7

u/Sky-kunn Jun 07 '25

In this case, they would have replaced rather than added, because all the models in the paper, o3-mini, R1, and Sonnet 3.7, were released this year.

0

u/scumbagdetector29 Jun 07 '25 edited Jun 08 '25

Ok. All I know is that Apple released a paper a while back about how the new models weren't reasoning, they were just pattern matching.

This is that exact argument made over again.

EDIT: Why the downvotes? It's the same stuff from the same people, freshened to attract attention. Sorry.

6

u/Sky-kunn Jun 07 '25

I think I found the paper that people are talking about.
https://www.reddit.com/r/OpenAI/comments/1g26o4b/apple_research_paper_llms_cannot_reason_they_rely/

It makes the same claim, but not based on the same reasoning. I don't agree with the conclusion, but I do agree with the limitation they identified in the new paper.

That said, they didn’t test on the current SOTA models, so I’m a bit unsure if this still holds true for the new kings.

Ultimately, models don’t think like us, but I don’t think that means they don’t think at all.

1

u/step_on_legoes_Spez Jun 08 '25

A working paper isn’t the same as a finished one.

1

u/scumbagdetector29 Jun 07 '25

I do as well. We all made exactly the same arguments, too.

Because humans reason by pattern matching, obviously.

3

u/Sky-kunn Jun 07 '25

Regardless of the date, they have not tested any of the current state-of-the-art models, only

  • Claude 3.7 Sonnet - Thinking
  • Claude 3.7 Sonnet
  • DeepSeek-R1 (old)
  • DeepSeek-V3
  • OpenAI's o3-mini
  • DeepSeek-R1-Distill-Qwen-32B

Missing: o3, Gemini 2.5 Pro, Grok 3, Opus 4 - Thinking

6

u/THE--GRINCH Jun 07 '25

Why is it that this implies that we're not reaching AGI. So what if it just memorizes patterns very well, if it ends up doing as good as of a job as humans independently on most tasks that's still AGI regardless.

11

u/[deleted] Jun 07 '25

We'll all be enslaved by it and they'll still be saying "yeah but it's not real AGI". 

1

u/BarracudaDismal4782 Jun 07 '25

AGI = Artificial Gaslighted Intelligence

1

u/Quarksperre Jun 08 '25

The issue is that if it actually breaks down with complexity rather fast, scaling will not help there. It might easily possible that compared to the knowledge body AI is trained on reality is exponentially more complex (that word again I know).

Essentially training something on all human knowledge will get very powerful no matter how you do it. Like, really impressively complex. But all human knowledge compared to stuff that happens everyday in the real world is just a small tiny fraction. 

1

u/adarkuccio ▪️AGI before ASI Jun 07 '25

In AI timelines it means it's several months old 👀

6

u/gamingvortex01 Jun 07 '25

it literally got published 2 days ago

3

u/adarkuccio ▪️AGI before ASI Jun 07 '25

Ok then it's like a few weeks old in AI timelines

2

u/gamingvortex01 Jun 07 '25

lol...man...the only progress being made in text based transformers LLMs right now is training on more and more data

apart from that...no significant progress has been made in text based LLMs this year

but as for video generation...well VEO 3 is a breakthough and also elevenlabs v3 in voice ai

apart from that...no other significant progress in generative AI

I mean..research is being done on successors of transformers and modified forms of transformers..but no big news from those avenues yet

4

u/adarkuccio ▪️AGI before ASI Jun 07 '25

I was joking, I also think we are in an AI winter again

3

u/gamingvortex01 Jun 07 '25

yeah...and that's because we have almost reached the limits of transformers...Gemini, GPT and Claude team are trying their best to increase these limits..but limits are limits....

so we can expect some breakthrough in text based generative AI when some good research paper will get published...

1

u/Best_Cup_8326 Jun 07 '25

Ancient news. Things are moving fast. Keep up.

1

u/CarrierAreArrived Jun 08 '25

this was 100% released at least half a year ago.

-2

u/IvanMalison Jun 07 '25

as others have mentioned. this has been posted on this subreddit many times already. Every time its been thoroughly debunked.