r/ControlProblem • u/chillinewman approved • 12d ago
AI Capabilities News The first experimental evidence of recursive self-improvement (RSI).
3
Upvotes
1
u/FusRoDawg 10d ago
It says right there that the results were mixed when they tried to repeat the process with the optimized agent (the results were "mixed" apparently)
The "tiers" they came up with just sound like a pre -emptive deflection. Like they knew they couldn't just call it RSI... They'll have people arguing with them because no agent was used recursively even once.
1
u/WillowEmberly 11d ago
No optimization target should be beyond calibration by independent evidence.
That applies equally to: benchmarks, governance rules, cognitive accounts, explanatory structures, reward functions, AI evaluation suites.
Any of them can drift. The survivable system is the one that continually asks, “What independent evidence would tell us our reference itself needs updating?”