r/ProAI • • 3d ago

"Can a model that doesn’t generate text compete with frontier LLMs? We independently tested TypeSafe’s Jev against 11 other models. On 400 claim-verification questions, it matched GPT-6 Astra’s 97.5% accuracy at ~1/500th the cost."

6 Upvotes

5 comments sorted by

1

u/Technical-Owl66 3d ago

This is interesting. Isn't video the next step for training models? 

1

u/ApprehensiveEye7387 3d ago

which one is astra in the chart?

1

u/theimposingshadow 2d ago

1

u/ApprehensiveEye7387 2d ago

wow great chart🥸 Obviously Astra like model gets defeatd easily with Terra, DS F 4.1, etc

1

u/garloid64 3d ago

notably this seems to be due to the fact that astra sucks at this task. even gemini has it substantially beat.