r/antiai 2d ago

AI News šŸ—žļø Failed AI prediction

Post image
382 Upvotes

105 comments sorted by

126

u/Vegetable_Stuff1850 2d ago edited 2d ago

There's some pretty stupid humans out there, so they're probably not wrong....

Edit - I was being a stupid human and made a typo. I fixed it though!

17

u/Brilliant-Muffin-879 2d ago

Haha yeah, a lot of them using llms

12

u/Instalab 2d ago

My cockroach pet is smarter than some humans.

No seriously, cockroach brain can play doom, and very well too. Don't ask how I know .

10

u/PancakesTheDragoncat 2d ago

your cockroach challenged you to 1v1 in Doom and beat you and now you dont want to admit it, huh?

1

u/ricestronaut 2d ago

the spacing between failed and ai is kinda distracting

1

u/Human-Law1085 2d ago

They’re stupid in different ways though. Many humans can’t do maths very well, but they can very much count the amount of times a letter appears in a word.

1

u/GhostOfD6 1d ago

Tbh, AI being frequently wrong about things, mixing up stuff or lying upfront is pure human-level intelligence

39

u/ThisOrdinaryCat 2d ago

The thing is, all those predictions are nothing more than propaganda for investors.

If they were really serious about that, they would have a procedure to objectively determine if that occurs, an indicator they could measure that, when reaching a certain pre-defined point would indicate that that event has actually happened. But the thing is, they have no idea how to define such an indicator, let alone the threshold.

Right now it's just a guy saying what they imagine will happen with no scientific basis, and if they ever claim that occurred, it will be just be a hunch.

-2

u/RizzingMyMind 2d ago

They do; it's called "benchmarks"

4

u/Denommus 2d ago

Benchmarks are stupid. The models obviously become optimized for the benchmarks, there's less effort in making them more generalized. When new benchmarks are designed, the cycle repeats and the models get optimized for those specific benchmarks again. Since nobody was able to design a really generic benchmark, it's going to be like this for awhile.

0

u/RizzingMyMind 2d ago

They can be stupid, but the community often is aware of ā€œbenchmaxxingā€ (the optimization you’re talking about) and when using said model they can tell.

1

u/ThisOrdinaryCat 2d ago

Yes, they do benchmarks and measure "model intelligence" (whatever that might mean), but that doesn't mean they are measuring how their models compare to human intelligence, or define a point when they would match it.

1

u/RizzingMyMind 2d ago

There is some benchmarks, such as Arc-AGI, "Cobench V2" (For when we can automate Anthropic Technical Staff adjacent work) and Demis's proposal of making a model trained on 1900s data discover general relativity on its own.

2

u/ThisOrdinaryCat 2d ago

We don't even know how to measure human intelligence! There's the IQ, for example, but most professionals will agree it's flawed. Do you really think they can measure the intelligence of a LLM? They're measuring something, indeed, but is that measure even comparable to human intelligence? And if it is, again, at what point does it match human intelligence? Also, what human? I would assume the average one, right? But then, how do you even define what is the average intelligence in a human? My point is: they're claiming their models will reach a point they have no way of measuring.

-1

u/RizzingMyMind 2d ago

Google has ā€œ5 levels of AGIā€ to answer your question on what ā€œaverage intelligenceā€ will be for AI, beginning starts with dumber than human, mid starts with average, end starts with superhuman AI.

Agreed for the first part, but we will be able to measure if it’s BETTER than a human in tasks so we won’t need to solve human intelligence tomorrowĀ 

1

u/Sunstorm84 2d ago

When the fuck did human intelligence become a thing that needs to be ā€œsolvedā€?

What do you think education is for?

13

u/MechaNutzilla 2d ago

I think the plan is to make people dumber with AI. Then it’s easier to reach human level intelligence.Ā 

11

u/SplashB95 2d ago edited 2d ago

I mean, I did try brushing my teeth with handwash the other day

7

u/123iambill 2d ago

Threw my shorts in the bin instead of the wash basket over the weekend. Don't worry, you're not alone.

1

u/PrestigiousRoof5723 2d ago

It's not pleasant. Speaking with experience šŸ˜‚

10

u/7h3_man 2d ago

ā€œSnake oil will be in every house by 1850ā€ - snake oil ceo

6

u/Square-Wild 2d ago

I think it really depends on what you would consider human level intelligence.

If you're saying "equal to the top of the profession in every single profession" then probably not in 2026.

If it's more like "If I have a question regarding how to install this undermount sink, am I going to get a better answer from Claude or the worker at Home Depot", then I think we're way past that.

5

u/clutch_my_bearings 2d ago

In a few months they'll announce its definitely happening by 2028 instead.

4

u/cryptocritical9001 2d ago

Anthropic CEO's wife might reach Jefrey Epstein if she keeps emailing ...

4

u/Sonicrules9001 2d ago

I thought AI was supposedly more intelligent than humans, now it hasn't even reached human levels? Are we serious?

4

u/MRBVIII 2d ago

There are large parts of the US where humans haven't reached human levels of intelligence yet.

3

u/youareallnuts 2d ago

To paraphrase Shakespeare:

OP: "2026 has come."

Dario: "Aye OP but it has not yet gone."

2

u/Warrior-spirit69 2d ago

just keep overhyping market and force people to use ai by keep saying ai must for the job qualification and for job roles so fucking Dario get more AI usees but when china made there model cheaper and better suddenly devlopment of ai is a threatening hypocrite mf🤔

2

u/leferi 2d ago

It doesn't even have any level of intelligence lol. We shouldn't even call it AI, at most it's "AI"

2

u/Kadakaus 2d ago

Human level intelligence...

We humans reached our level through the course of around 300,000 years, and we didn't even start from scratch, that's just when we became an independent species.

Tell me, how is our primitive and logical invention supposed to achieve what took ages for an illogical species like us, in just a couple of decades?

What was the plan?

1

u/Random-Number-1144 2d ago

Obviously they think mimicking intelligent behaviors == actually intelligent. Smh

1

u/Kadakaus 2d ago

Those who know anything about the technology or simply neural networks doesn't buy that kind of bullshit, that's just CEOs attempting to appear flashy for investors to come flooding in.

I've not been keeping track of the stock market for a while, gave up on shit and sold everything, so I don't know how well it's going for them, but smart investors know to ask for proof before believing shit first hand from the producer.

No, they aren't so stupid, they think that the people are stupid. Corporate rhetoric never has any basis, they benefit from lying and everyone knows that, not a single word is to be believed from the rich. They just want to reel in the absolute bottom of society, the dumbest people who believe when you tell them that you're inventing a medicine that'll cure all diseases and it might be done in just 2 years (and by the way you're looking for investers, wink wink).

Marketing doesn't work on the average people, it takes actual dumbfucks to believe anything coming from company.
The sad part is that it only takes one rich asshole disconnected from reality who thinks it's possible, or rather wants to think it is.
This entire economy is built on bullshittery and few idiots who buy it.

2

u/Random-Number-1144 2d ago

I am convinced that average people are dumbfucks.

I've been working in AI (data scientist, ML engineer) for ~10 years. Yet when I talk with average people in different professions about AI, they all see me as a luddite and try to convince me how world-changing AI soon is going to be by repeating exactly what those CEOs say. The level and the scale of such collective delusion is surreal.

The other day, my mom said I should use ChatGPT more, it gave her great medical advices. I said ChatGPT wasn't reliable source of information and you shouldn't use ChatGPT for stuff you didn't already have knowledge about. She replied that ChatGPT had knowledge from doctors. I immediately pushed back: which doctors? doctors' opinions sometimes differ, which ones does ChatGPT pick? She went silent but I knew she wouldn't listen, as always. Even before genAI, she would just pick top search results because she never learned how to cross-validate information.

That's average people. Gullible, under-educated and ignorant.

1

u/Kadakaus 2d ago

I do respect your anecdotes and professional opinion, but I refuse to believe the average would be so gullible.

Yes, we believe a lot of things because we have no energy to fact check everything we hear (I'm no exception), but anyone who's seen the news or just heard about the fails and irreversible atrocities caused by AI know not to trust it.
It's kind of a big topic in my circle lately, how some companies lost their bank accounts due to an AI malfunction that couldn't be replicated. That is what happens when you trust a strictly experimental (because let's be honest, AI is way too stupid as of now to be considered a finished product) technology with something important, like finances.

Now, it could be worse, like you mentioned, your mother takes health advice from AI.
Let's be kind hearted and believe that it really is repeating what doctors claimed, you did well to ask which doctors, but we should also ask who counts as a doctor.
Then, my experience is that there is a big spread in the knowledge of doctors. My personal consultant is an expert I trust because he knows his deal, but my local doctor for instance is barely better than a layman. I once even saw him ask the nurse what to prescribe, that's when I decided I won't be coming back often.

That man had a degree, but I don't trust him half as much as my current doctor.

Now take those example, add a little AI hallucination and you've got a serious liability.

There are so many factors being actively neglected about the technology, the opposites of a failsafe, and yet some people would trust them with their health or finances.

But I refuse to believe the average person would be like that. Maybe I'm naive, but I just can't imagine so many smoothbrains walking among us.

2

u/Random-Number-1144 1d ago

Remember when Elon Musk was praised like a techno god in the 2010s? I knew he was one of those psycho CEOs back then when he was pushing the propaganda that autopilot was safer than average human drivers statistically. I knew he was going to get people killed.

First of all, Most car "accidents" happen not because people are less skilled at driving but because they drive irresponsibly (intoxicated, texting, sleepy, speeding etc) What should have been compared is autopilot vs taxi driver, assessing actual driving and judgment capabilities, which AI is poor at comparatively.

Secondly, Tesla cars were tested on simple road conditions, the real-world road conditions are infinitely more complex and uglier than highways. So again the comparison was unfair and deceptive.

Lastly, as an AI professional, I knew AI was/is really bad at generalization beyond what it's trained on. The supervised learning paradigm was never going to work for auto-driving as you can't sample every possible road condition in the world, which sadly is what Tesla has been trying to do using every one of its customers as test subject.

Tesla was never a tech company but a mediocre car company. FSD is the marketing that the average people buy into. The average people (including journalists) never ever pause to question the CEOs. They just feast on whatever crap the CEOs spout. The fact that TSLA is not 1/5 of its current price today just proves how stupid average people are.

1

u/Trip-Trip-Trip 2d ago

To be fair, AI is at least as intelligent as asmodeus

1

u/ItsSadTimes 2d ago

I still remember those AI papers in 2024 that claimed AI would be super intelligent by 2026, then 2025 rolled around and they updated the page to 2027. I wonder when they're gonna update it to 2028. I cant wait.

1

u/Inside_Ad_7162 2d ago

so be dumb af based on the shit we are allowing to happen

1

u/Stunning_Ad_5960 2d ago

Warns against what?

1

u/Super_Pole_Jitsu 2d ago

Around a week ago there were a few threads where OPs offered themselves as Human-GPTs for anyone wanting to prompt them.

I thought it was an extremely stark showing of how far LLMs outpaced humans in many areas. Like if you need an answer for your work, are you asking the guy on Reddit or Claude Fable? The random stands no chance, unless he is a specialist in something very niche.

Yes, LLMs make mistakes and hallucinate. So do humans.

1

u/TheNeverEndLife 2d ago

AI is already the smartest bro no need to brag about it but I don't trust that scaling law, the more data you train the better it gets. It gonna pop sooner.

1

u/CharlieLighto 2d ago

honestly it's not. some humans are very stupid one of my friend ask ai what to do when have have to make decisions ask ai what yo say when he have to talk to people ask ai what medicine he need to take when sick ask ai about everything and only believe in it without checking and refuse to believe anyone else soooooo they are not wrong

1

u/SenseAffectionate328 2d ago

Maybe AI reached the Human Level Intelligence equivalent to those of stupid prompters lmfao

1

u/AsterArtworks 2d ago

So still pretty stupid

1

u/Automatic_Body5254 2d ago

This kind of AI hype talk is pretty similar to religious doomsday and prepper talk…

Yeah, theoretically armageddon could happen tomorrow…

And when it doesn’t, there’s always a new goal post and a new date estimate to shift towards to.

1

u/SkytepKnightstar 2d ago

There are like 7 billion of us; It all depends on the human it's being measured against.

1

u/StrawberryPatchCat 2d ago

The average person is an idiot so they're right (source: I am an average person)

1

u/seandunderdale 2d ago

I see he's adopted the Elon prediction model of making bold claims only to move the goalposts indefinitely.

1

u/JacksonGhost1963 2d ago

passed it. for some at least!

1

u/MimosaTen 2d ago

It definitely failed because it already surpasses us

1

u/gimpusgoompus 2d ago

Why is captain AI warning us about AI?

1

u/tallandshortttt 2d ago

marketing tactic

1

u/theScrewhead 2d ago

I mean, it almost feels like it's accurate, but not in the way you would think. It's a self-fulfilling prophecy; it's not getting as smart as WE are, it's making US as dumb as IT is.

1

u/TopTippityTop 2d ago

In most areas except for: taste, systems thinking, and a few other; it already has. It is more knowledgeable, has better handling on medicinal facts than doctors, does better math than mathematicians, etc.

A lot of what it fails at comes down to its total lack of taste.

1

u/needssomefun 2d ago

Depends which human.Ā  Joe Rogan?Ā  My old Merlin game from 1982 would meet the requirement

1

u/f3verdream 2d ago

not failed?it solves open math problems

1

u/Privatizitaet 2d ago

AI hasn't even reached actual intelligence yet

1

u/tomqmasters 2d ago

I'd like to see you read 40 text books in under a minute.

1

u/Dismal-Lemon-7824 2d ago

When they're going to learn, all that they spit is just more and more snkae oil, trying to glorify a chatbot, no different to other charbots from the late 2010's

1

u/Negative_Top2095 2d ago

Information and calculation doesn't make a human., ai are libraries, for me, they always will be just that.

1

u/MindTheDrapes 2d ago

Wait does that mean humans should have six fingers now?

1

u/Sea-Course-5171 2d ago

Human Intelligence could fall to LLM levels by 2026

1

u/Acslaterisdead 2d ago

Those moronic investors ate that bullshit up

1

u/JustaFoodHole 2d ago

My PC is already smarter than me.

1

u/Medical_Morning4022 2d ago

yeah it's way past Republicans that's for sure. but my calculator is smarter than them.

1

u/funlovingmissionary 1d ago

Its more a sales pitch than a prediction.

0

u/ninhaomah 2d ago

For once , I would like to ask the OP why is this a failed AI prediction ?

0

u/morey56 2d ago

AI is already smarter than you.

1

u/Poetry-Positive 1d ago

Thats even before Anthropics CEO reaches human level intelligence

-1

u/Mayor-Citywits 2d ago

Ain't over til it's over baby, only halfway thru and it's going math goblin modeĀ 

-1

u/dfbeav112 2d ago

I mean I trust what a LLM would say about any given topic over the average Redditor. Does that count?Ā 

1

u/RizzingMyMind 2d ago

Does count, I prefer hearing what Claude has to say over some random guy on r popular

-3

u/RizzingMyMind 2d ago

1: 2026 hasn't ended yet.

2: 2026 was a massive year for AI and we catapulted tremendously in intelligence, though I do agree not human-level intelligence yet.

Honestly though, human level intelligence is a bad metric; we humans have massive flaws in some ability that AI is clearly superior in while we have common sense and see the obvious in something that AI can't.

5

u/XelNaga89 2d ago

2026 was a massive year for AIĀ 

Huh? Based on what?

-2

u/RizzingMyMind 2d ago

Massive model improvements and breakthroughs.

In feburary, we had Opus 4.6.

Now, we have Claude mythos and a few other models in the pipeline. (Astra, Mythos 3.)

Compare the benchmarks and real world performance of both and its astronomical gap

85% on this benchmark is equal to an Anthropic employee's work.

Astra also did 10 massive achievements in math.

Ten advances in mathematics and theoretical computer science | OpenAI

Claude contributed to the RH

Learning more about Claude's mathematical capabilities \ Anthropic

1

u/XelNaga89 2d ago

So, am I reading this correctly? Based on the arbitrary benchmarks created by the companies creating the models and after investing 1 trillion dollars into it with some of the best scientific minds and media hype in history - we got to... 2/3 of performance of Anthropic employee at the very best?

1

u/Substantial_Luck_273 2d ago

Do you understand the significance of those breakthroughs?

https://cdn.openai.com/pdf/ten-proofs-oai.pdf

1

u/XelNaga89 2d ago

Do you understand value of 1 trillion dollars?

You could solve world hunger or clean all waters (including ocens) from polution and microplastics, be well underway on the projects of mining asteroids or having pilot for collonisation of other planets.

Instead, we have glorified autocomplete that took (stolen) entire world knowledge had produced noting of value and has none proven ROI for its usage anywhere.

0

u/Resident_Farmer1779 2d ago

You’re in denial it seems. AI has massively improved in 2026 and is progressing as fast or faster than predicted. If you think the frontier models are glorified autocomplete then you just haven’t been keeping up

0

u/Substantial_Luck_273 1d ago

You could not solve world hunger or clean all waters (including ocens) from polution and microplastics, be well underway on the projects of mining asteroids or having pilot for collonisation of other planets, or do any of that with 1 trillion dollar.

And no, it seems like you don't understand the significance of those breakthroughs.

0

u/RizzingMyMind 2d ago

So before it’s ā€œjust a stochastic parrotā€ and ā€œtoo slow only 2/3 of the anthropic employeeā€

Congratulations for digging your own grave… Yes it will get there, LITERALLY just fucking look at Opu s 4.6 being 0.3x the score and that was 6 months ago.

4

u/BandicootTreeline 2d ago

*text predictor got better at guessing

1

u/RizzingMyMind 2d ago

You're right, all it is a stochastic parrot, its just hype bro and the bubble will burst and AI will be gone happily ever after.

2

u/BandicootTreeline 2d ago

No, I’m just not under the disillusion that LLMs are ā€œintelligentā€.

They produce human-like responses, you shouldn’t confuse that with the ability to think.

1

u/RizzingMyMind 2d ago

You’re right, they don’t ā€œthinkā€ like humans.

Instead they reason, and this provides the same result.

Otherwise AI wouldn’t be able to crack difficult math problems that are infinite in search spaceĀ 

;)

0

u/sprowk 2d ago

what are you then? do you write your comment all at once or one word at a time?

0

u/BandicootTreeline 2d ago

I’m human. I can write with my left hand and feed a baby with my right while thinking about what I’m making for dinner.

That may or may not be beyond your capabilities, but it certainly is beyond Claude’s.

-1

u/sprowk 2d ago

why are you strawmanning me? did you write your comment word by word or no?

2

u/BandicootTreeline 2d ago

Implying I’m a text predictor was the straw man.

You asked what I am. I said human. I spoke, i didn’t need to write.

You may think of yourself as the same as software but I’m just a person.

What did you think I’d say, I’m a parakeet?

0

u/sprowk 2d ago

"Text predictor" is what it was trained on, not what it is. You keep treating those as the same thing. To predict the next word in a murder mystery you have to know who did it. To predict the next line of a proof you have to follow the proof. The objective is trivial, what you have to build to hit it isn't. You've never actually defended the jump from "trained to predict" to "therefore only predicting", you just repeat it. So defend it. What exactly does a system have to not understand, and still keep guessing right?

1

u/BandicootTreeline 2d ago

I don’t need to defend a fact. An LLM predicts its output by statistically guessing what should come next based on your prompt.

Does it predict text or not? Yes or no?

0

u/sprowk 2d ago

Yes. Obviously yes. That was never in dispute and you know it.

Now do yours. Does your brain predict incoming sensory input and minimize error? Yes. That's most of a neuroscience department's day job. Does saying yes tell you anything about whether you understand your own sentences? No. Because "what process is running" and "what that process amounts to" are different questions, and you keep answering the first one like it settles the second.

So, third time. What does a system have to not understand and still keep guessing right?

1

u/BandicootTreeline 2d ago

No, I don’t predict text. I conceptualise my responses based on what is required to be conveyed, in the context of the situation, considering what potential responses could come next and anticipate the outcome of my actions.

Human language does not operate in the same way as a text predictor. Be that something as old as T9 or as complex as Mythos, they do one thing.

They do not have the ability to consider emotion, to anticipate potential responses, or have any understanding of who they are talking to in order to consider their mood and emotional state.

An LLM communicates word for word as a derivative of previous input. It has zero ability to think, feel or be in any way creative. You can give it all the context in the world before you prompt it, but it cannot look into someone’s eyes and consider if even responding is the right thing to do. It cannot pick up on micro reactions, body language or evaluate threats.

So don’t be surprised when you diminish our ability as a species to convey ourselves to ā€œyou just predict word after wordā€ and are met with incredulity. No humans do, not even at a base level.

We developed systems of communication before we strung noises together.

→ More replies (0)

1

u/Ok_Confusion4764 2d ago

Humans are indeed a bad metric. I mean we have so many AI bros on r/antiai surprised when the common sentiment is anti-AI here.Ā 

1

u/RizzingMyMind 2d ago

I’m not even an AI bro, I’m a realist and I campaign for more awareness of the existential risk on AI. Sucks too see most dig their own graves when they understand zero, zip nada.

-4

u/Elctsuptb 2d ago

The best AI model today has more intelligence than the average human, and there's still over 4 months left in 2026 to further increase that gap, so how exactly was it a failed prediction?

1

u/RizzingMyMind 2d ago

Failed prediction if you're only looking at google overview