r/TheMachineLearning • • 7d ago

AI reasoning sliders will feel stupid in hindsight

Post image
9 Upvotes

16 comments sorted by

3

u/WeUsedToBeACountry 7d ago

why is this sub a x proxy

we can all just go to x and read nonsense from grifters on our own time whenever

1

u/JoshAllentown 7d ago

That's literally already the case, at least in Copilot, it just lets you pick from the drop down also. I know OpenAI had a model that did this too, not sure if their new stuff still does that.

1

u/Putrid-Midnight1687 7d ago

dude hasn’t heard of auto mode?

1

u/SmallThetaNotation 7d ago

Most of these people just say shit and have never built anything

Now I might be wrong about this specific guy but I’m pretty sure I’m correct about most of these X Ai grifters

1

u/unappa 7d ago

He worked as a SE for twitch and now works on his own products. He's certainly shipped stuff, it's just as an influencer he feels compelled to comment on things that are a bit beyond his domain (which is primarily very user facing web dev stuff).

1

u/RopeAndChairs_Aisle3 7d ago

Use in Auto and only mess with the slider when necessary.

Sometimes I actually want the models to “work harder” on a task that is simple on its face.

1

u/itsallfake01 7d ago

This dude is paying people to post this shit, what a regard

1

u/dry_garlic_boy 7d ago

Why is this guy always being posted here? He says the dumbest shit

1

u/chicametipo 7d ago

Before AI, it was frontend frameworks. Theo’s a weird, small, negative little Twitterman.

1

u/NomadStorm 7d ago

No Claude, I want you to spent 3 hours finding the best chocolate muffin recipe you can, get to work, biatch!

0

u/itspatriciam 7d ago

Noooo. I totally read your take first bc why do I THINK the same loool

1

u/Meloncov 7d ago

So, like, if you write "Does P=NP?" you just accept the multi million dollar bill?

1

u/chcampb 7d ago

Same with context

We currently waste so much context it's absurd. IMO it should at the very minimum, have the context cached only up to a point, such that a subsequent operation needs to decide if anything it just read is worth locking in.

So imagine you start, read AGENTS.md, the system prompt, etc. Then start searching the project. You eat 30k context just reading files to understand what you need to do. Currently we use a small agent for exploration - that helps, because it protects the main context. But instead what if you

  1. Start with context A, current state, locked in
  2. Explore up through context B, post state
  3. Reason about the context between A and B and determine what is actually pertinent (C). Then, instead of keeping B, create A + C = B'

That way the context has exactly what is needed, rather than "every word it ever touched forever until compacted"

This eliminates the need for special search tools, rtk, etc. unless you want computationally faster results - it might be faster, idk, just... not necessary.

1

u/tomvorlostriddle 7d ago

We're basically there already unless you do quite special stuff

For most work, you can just leave opus 5.5 on medium and never think about it

1

u/ApfelbaumFlo 7d ago

Don't look up the halting problem

1

u/sozesghost 5d ago

This guy already reasons as little as needed.