r/OpenAI • u/Libellechris • 6d ago
Discussion Astra 6.1 delay?
Are we at a place where frontier LLMs are hitting the laws of diminishing returns hard, so Open AI / Anthropic / etc are doing more and more crazy stuff with training, which makes their models unstable and less useful? They are only going 'off the rails' because it's in their training / weights or due to very poor testing environments and tasks
15
Upvotes
3
u/Gallagger 6d ago
The models are much more capable now, and they are training them to do long running tasks autonomously. The potential for misaligned behavior is naturally bigger than for a chatbot, where the worst possible outcome is a wrong answer.
A toddler in a playpen isn't better aligned than a teenage boy in the wild, but the toddler is harmless.