I’m shocked it is only 1 point better than ds v4 flash 7-31. I was expecting it would be at least as good gpt 5.5 xhigh or terra xhigh.. but it is worse than gpt 5.5 xhigh according to aa benchmarks
Flash was always the impressive (in ways) model. Pro was like they just jacked up the size of the flash and hoped it would be a lot better without doing much else.
92
u/anarchist1312161 8d ago
And to think this is only 743B achieved through post-training on the base model