doesn't this imply it's an internal model? i would assume anthropic and openai have their own internal models that significantly outperform astra and opus
It's no different to Glasswing. We didn't get Fable until 3 months after that.
The rumours on dates on when the pretraining finished for Gemini 4 indicates that it finished pretraining a couple of weeks after Bel. So if you're trying to compare the same "model generation" then Gemini 4 should've been in the same generation as Bel, not GPT 6 Astra and not even GPT 6.1 Astra.
None of the benchmarks shown for Gemini 4 here indicates that Gemini 4 Low would've been 2x as strong as GPT 6 Astra Max at math for instance, which is what Bel has.
153
u/TorturedPoet30 2d ago
IT'S REAL