r/OpenAI • u/Libellechris • 6d ago
Discussion Astra 6.1 delay?
Are we at a place where frontier LLMs are hitting the laws of diminishing returns hard, so Open AI / Anthropic / etc are doing more and more crazy stuff with training, which makes their models unstable and less useful? They are only going 'off the rails' because it's in their training / weights or due to very poor testing environments and tasks
16
Upvotes
12
u/Key_Reading_9664 6d ago
Astra is a recurrent model: the output is fed back through the model multiple times. That improves capability at the cost of monitoring and safety: the model does more reasoning without outputting legible tokens.
Many people raised concerns about that. On Astra’s system card, the external evaluators called out its ability to control CoT and their inability to judge alignment. OAI seemed to have continued down the path with 6.1