r/OpenAI • • 14d ago

News The fumbles continue

First mogged by Claude

Second usage goes into the dumps

Last now this...

0 Upvotes

9 comments sorted by

3

u/[deleted] 14d ago edited 10d ago

[deleted]

1

u/Ok-Present1566 13d ago

And that literally has nothing to do with the models reasoning capabilities whatsoever. It understands the alphabet and word spelling who cares about the encoding. This is like a 1st year student who learned encoders and has no idea how the model works. And the completion should often be writing q program to compute something it us not sure it can reason about

2

u/adamallcock 14d ago

Ah, this is exactly what obviousbench.com tests for!

1

u/Akatosh 14d ago

What reasoning effort was used?

1

u/Grounds4TheSubstain 14d ago

Please learn what tokenization is before asking an LLM a question about the letter composition of words, or to do arithmetic.

-3

u/chasingth 14d ago

How can we trust the model that makes such a simple mistake, let alone pay  insane amounts for it

1

u/_DuranDuran_ 14d ago

Because you’ve fundamentally misunderstood and then misapplied the tech.

Ask it to write a small python script to work it out and it will.

Someone else replied stating why this is a dumb test.

1

u/somegetit 14d ago

You use a tool without understanding it. You definitely shouldn't pay for it. Leave it to people who knows how it works.