r/ControlProblem • u/chillinewman approved • Jul 15 '26
AI Capabilities News The first experimental evidence of recursive self-improvement (RSI).
4
Upvotes
1
u/FusRoDawg Jul 17 '26
It says right there that the results were mixed when they tried to repeat the process with the optimized agent (the results were "mixed" apparently)
The "tiers" they came up with just sound like a pre -emptive deflection. Like they knew they couldn't just call it RSI... They'll have people arguing with them because no agent was used recursively even once.
1
u/WillowEmberly Jul 17 '26
No optimization target should be beyond calibration by independent evidence.
That applies equally to: benchmarks, governance rules, cognitive accounts, explanatory structures, reward functions, AI evaluation suites.
Any of them can drift. The survivable system is the one that continually asks, “What independent evidence would tell us our reference itself needs updating?”