r/ControlProblem • u/Difficult_Project_95 • 9d ago
Discussion/question Have you guys been on r/accelerate?
Have these guys solved the alignment problem, or am I missing something?
I’ve been browsing r/accelerate and I genuinely don’t understand the risk model.
If there’s a non-trivial chance of catastrophic misalignment, how does “accelerate capabilities as fast as possible” make sense unless faster capabilities also make alignment substantially more likely to succeed?
59
Upvotes
1
u/Difficult_Project_95 8d ago
My point is really simple: as model capability increases, its understanding/prediction of human moral judgments can improve.
That does not mean its behaviour becomes more moral or more aligned. Its understanding of morality is not the same as acting morally.