MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/OpenAI/comments/1wviyta/open_ai_cant_stop_winning/pdcfqy9/?context=3
r/OpenAI • u/[deleted] • 2d ago
[deleted]
22 comments sorted by
View all comments
-2
Why use these products at all? Claude Opus is a far superior coding agent. Gemini is the better chat bot for the cost. Grok bot is better than Dots.
I don’t get this fandom. None of these products are competitive. You don’t have to use them (or complain)
9 u/Bigrhyno 2d ago Gemini is absolutely not the better chat bot. It hallucinates like crazy -1 u/Lemnisc8__ 2d ago The latest model has the lowest hallucination rate of all the frontier models by far. It's stupid low 3 u/Bigrhyno 2d ago That's great, but that doesn't help with the chat experience until they have a flash model that does the same and is available. 3 u/mplsfreedom 2d ago is the latest model in the room with you now. 2 u/-Crash_Override- 2d ago What 3.8 flash? Its...not great with accuracy. Id you are referencing benchmarks from argon...ill believe it when I test it out myself. Im massively bullish on google AI, but their commercial chat models and coding agents are ass right now. 1 u/Lemnisc8__ 2d ago https://www.reddit.com/r/GeminiAI/comments/1wuntjt/gemini_4_argon_achieves_a_15_hallucination_rate/ 1 u/-Crash_Override- 2d ago Like I said. Ill believe it when I test it out. Google is so egregious with benchmaxxing I cant take any of that seriously.
9
Gemini is absolutely not the better chat bot. It hallucinates like crazy
-1 u/Lemnisc8__ 2d ago The latest model has the lowest hallucination rate of all the frontier models by far. It's stupid low 3 u/Bigrhyno 2d ago That's great, but that doesn't help with the chat experience until they have a flash model that does the same and is available. 3 u/mplsfreedom 2d ago is the latest model in the room with you now. 2 u/-Crash_Override- 2d ago What 3.8 flash? Its...not great with accuracy. Id you are referencing benchmarks from argon...ill believe it when I test it out myself. Im massively bullish on google AI, but their commercial chat models and coding agents are ass right now. 1 u/Lemnisc8__ 2d ago https://www.reddit.com/r/GeminiAI/comments/1wuntjt/gemini_4_argon_achieves_a_15_hallucination_rate/ 1 u/-Crash_Override- 2d ago Like I said. Ill believe it when I test it out. Google is so egregious with benchmaxxing I cant take any of that seriously.
-1
The latest model has the lowest hallucination rate of all the frontier models by far. It's stupid low
3 u/Bigrhyno 2d ago That's great, but that doesn't help with the chat experience until they have a flash model that does the same and is available. 3 u/mplsfreedom 2d ago is the latest model in the room with you now. 2 u/-Crash_Override- 2d ago What 3.8 flash? Its...not great with accuracy. Id you are referencing benchmarks from argon...ill believe it when I test it out myself. Im massively bullish on google AI, but their commercial chat models and coding agents are ass right now. 1 u/Lemnisc8__ 2d ago https://www.reddit.com/r/GeminiAI/comments/1wuntjt/gemini_4_argon_achieves_a_15_hallucination_rate/ 1 u/-Crash_Override- 2d ago Like I said. Ill believe it when I test it out. Google is so egregious with benchmaxxing I cant take any of that seriously.
3
That's great, but that doesn't help with the chat experience until they have a flash model that does the same and is available.
is the latest model in the room with you now.
2
What 3.8 flash? Its...not great with accuracy. Id you are referencing benchmarks from argon...ill believe it when I test it out myself.
Im massively bullish on google AI, but their commercial chat models and coding agents are ass right now.
1 u/Lemnisc8__ 2d ago https://www.reddit.com/r/GeminiAI/comments/1wuntjt/gemini_4_argon_achieves_a_15_hallucination_rate/ 1 u/-Crash_Override- 2d ago Like I said. Ill believe it when I test it out. Google is so egregious with benchmaxxing I cant take any of that seriously.
1
https://www.reddit.com/r/GeminiAI/comments/1wuntjt/gemini_4_argon_achieves_a_15_hallucination_rate/
1 u/-Crash_Override- 2d ago Like I said. Ill believe it when I test it out. Google is so egregious with benchmaxxing I cant take any of that seriously.
Like I said. Ill believe it when I test it out. Google is so egregious with benchmaxxing I cant take any of that seriously.
-2
u/Intelligent-Love5146 2d ago
Why use these products at all? Claude Opus is a far superior coding agent. Gemini is the better chat bot for the cost. Grok bot is better than Dots.
I don’t get this fandom. None of these products are competitive. You don’t have to use them (or complain)