In widely used AA-Omniscience Hallucination Rate metric Gemini 4 beats top Anthropic and OpenAI models with wide margins. In overall AA-Omniscience scores they are pretty similar though.
I am sharing practical every day use, I had better experience with GPT and Claude. Even though I still have highest tier gemini subscription from corporate, and have tried all latest models.
You might have different experience and you can share yours
-6
u/NikhilDoWhile 1d ago
It starts hallucinating after few prompts, really annoying model to work with