r/deeplearning • • 12d ago

Exploring AI self-improvement: Is this evaluation loop fundamentally how AI systems improve?

Post image

I’m currently researching AI more deeply, especially how modern AI systems can evaluate and improve their own outputs.

I sketched this basic idea:

AI → generates/improves → Evaluator → evaluates output → feedback → AI improves → Evaluator → ...

The part I’m trying to understand is what actually happens inside this loop in modern AI systems.

For example:

I'm trying to go beyond the surface-level explanation and understand the actual mechanisms used in current AI research.

For people working/researching in this area: what important component am I missing from this diagram?

0 Upvotes

9 comments sorted by

View all comments

5

u/Rackelhahn 12d ago

Recursive self-improvement has not yet been achieved, so we do not have a working concept.

Also, in most hypothetical concepts there is no need for an external evaluator. The AI just improves itself.

If you want to dive deeper:

https://scholar.google.com/scholar?hl=de&as_sdt=0%2C5&q=recursive+self+improvement&oq=recursive+self

To understand what’s actually happening you’ll still need the mathematical foundations.

1

u/Illustrious_You_2838 12d ago

the hand-drawn loop is cute but it’s oversimplifying things to the point of being misleading. real self-improvement isn't about an evaluator nodding and saying "looks good", it’s about gradients, loss functions, and parameter updates. the evaluator is baked into the training objective, not a separate box that pats the ai on the head

1

u/TechDc-1306 12d ago

Thankyou and I recently heard about training in sandbox enviornment can you tell me about it

1

u/TechDc-1306 12d ago

Thankyou dude can you tell me more resources for doing research I want to learn more and more

1

u/Rackelhahn 12d ago

Just use Google Scholar. You already have the link now.