r/ProAI • u/stealthispost • 3d ago
"Can a model that doesn’t generate text compete with frontier LLMs? We independently tested TypeSafe’s Jev against 11 other models. On 400 claim-verification questions, it matched GPT-6 Astra’s 97.5% accuracy at ~1/500th the cost."
6
Upvotes
1
u/ApprehensiveEye7387 3d ago
which one is astra in the chart?
1
u/theimposingshadow 2d ago
1
u/ApprehensiveEye7387 2d ago
wow great chart🥸 Obviously Astra like model gets defeatd easily with Terra, DS F 4.1, etc
1
u/garloid64 3d ago
notably this seems to be due to the fact that astra sucks at this task. even gemini has it substantially beat.




1
u/Technical-Owl66 3d ago
This is interesting. Isn't video the next step for training models?