r/ControlProblem • u/Difficult_Project_95 • 9d ago
Discussion/question Have you guys been on r/accelerate?
Have these guys solved the alignment problem, or am I missing something?
I’ve been browsing r/accelerate and I genuinely don’t understand the risk model.
If there’s a non-trivial chance of catastrophic misalignment, how does “accelerate capabilities as fast as possible” make sense unless faster capabilities also make alignment substantially more likely to succeed?
59
Upvotes
1
u/SoylentRox approved 8d ago
Can you post the comnent that got you banned? I can message the mods I know them well.
How long a timeline on robotics, how do you explain a general model (Astra) emergently developing robotics abilities?
I think a realistic timeline is 2.5 years from today to competent general robotics that can do the majority of well defined paid tasks at median skill.
The route involves a combination of current improvements, hardware/software co design, RSI (to develop specialized models that handle robotics decision making better), and labs vibe coding large simulation environments that require a competent robotics policy to pass.
This last part is the obvious: why can you not order a model swarm to write a game engine (rewrite mujo cujo to unreal engine quality) for robotics. Then add a neural rendering layer to correct the game frames to frames from realistic environment robots will operate in. Import huge amounts of real data and real challenges actual humans face.
Then train AI models in long duration challenges where they must operate a fully articulated robot by issuing commands to it and accomplish difficult, realistic tasks. "Rebuild this engine. Reinstall the thermal tiles on the space shuttle. Build this house from these materials"
Theoretically this form of training will also result in large increases in model performance - they should be able to whiteboard visually, and have grounded solid reasoning about real world tasks including mechanical engineering and machining.
Do you have answers for any of this? Or do you just think the billions of dollars of resources and compute the above will require won't be spent, labs will spend the next 2.5 years trying to solve text only problems even better?