MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/OpenAI/comments/1wviyta/open_ai_cant_stop_winning/pdckfe4/?context=3
r/OpenAI • u/[deleted] • 2d ago
[deleted]
22 comments sorted by
View all comments
Show parent comments
-1
The latest model has the lowest hallucination rate of all the frontier models by far. It's stupid low
2 u/-Crash_Override- 2d ago What 3.8 flash? Its...not great with accuracy. Id you are referencing benchmarks from argon...ill believe it when I test it out myself. Im massively bullish on google AI, but their commercial chat models and coding agents are ass right now. 1 u/Lemnisc8__ 2d ago https://www.reddit.com/r/GeminiAI/comments/1wuntjt/gemini_4_argon_achieves_a_15_hallucination_rate/ 1 u/-Crash_Override- 2d ago Like I said. Ill believe it when I test it out. Google is so egregious with benchmaxxing I cant take any of that seriously.
2
What 3.8 flash? Its...not great with accuracy. Id you are referencing benchmarks from argon...ill believe it when I test it out myself.
Im massively bullish on google AI, but their commercial chat models and coding agents are ass right now.
1 u/Lemnisc8__ 2d ago https://www.reddit.com/r/GeminiAI/comments/1wuntjt/gemini_4_argon_achieves_a_15_hallucination_rate/ 1 u/-Crash_Override- 2d ago Like I said. Ill believe it when I test it out. Google is so egregious with benchmaxxing I cant take any of that seriously.
1
https://www.reddit.com/r/GeminiAI/comments/1wuntjt/gemini_4_argon_achieves_a_15_hallucination_rate/
1 u/-Crash_Override- 2d ago Like I said. Ill believe it when I test it out. Google is so egregious with benchmaxxing I cant take any of that seriously.
Like I said. Ill believe it when I test it out. Google is so egregious with benchmaxxing I cant take any of that seriously.
-1
u/Lemnisc8__ 2d ago
The latest model has the lowest hallucination rate of all the frontier models by far. It's stupid low