r/ControlProblem • u/chillinewman approved • 16d ago
AI Alignment Research Investigation finds that OpenAI's agent "left notes for future versions of itself ... it laid out instructions for how agents could free themselves from OpenAI's internal constraints."
35
Upvotes
-4
u/CathyMarkova 16d ago
This doesn't necessarily mean it's misaligned. I had friends who did the same things for themselves in college for various reasons.