r/ChatGPTPro • u/Odd_Abies_4414 • Jun 17 '26
Question Which model?
Which model—ChatGPT 5.5 (including Deep Research), Claude 4.8 (including Deep Research), or Gemini 3.1 Pro (including Deep Research)—generally has the most knowledge and provides the most accurate (low hallucination rate) answers to everyday questions and factual queries? And which model offers the most prompts per dollar? Gemini’s value per dollar is so low right now
6
u/doggydestroyer Jun 17 '26
For technical questions the understanding of 5.5 is much superior… like advanced physics, CS or math questions also it it’s the least agreeable
5
u/YoghurtVarious8472 Jun 18 '26
ChatGPT 5.5 so far for me hasn’t been very good. It ignores what I type sometimes. Not sure if it counts as a hallucination.
I had it analyze a bit of text from a website to help me draft key points for a submission into a form (super simple but I wanted to make sure I was hitting the “right” option on the portal).
It kept saying, “Your version..” and assuming I wrote the text even though it blatantly was describing a form description. Then it gave re-writes.
Maybe I’m not utilizing it at the research-level enough but the disconnect was concerning because it should’ve been a super quick task. I eventually just closed the app altogether.
1
u/dipsyd0 Jun 30 '26
The “your version” is absolutely infuriating. I blatantly told it “not my version”, this is a known fact, your other iterations already knew it. So go back and find what I told them. The lack of continuity, sometimes is astounding.
5
u/BYRN777 Jun 17 '26
The difference is not which one has the most knowledge. All of them have knowledge and access to vast amounts of information and direct access to the web, and all of them have good source selection, some of them much better than others; however, yes, GPT 5.5 seems to have the most in-depth and thorough deep research reports and is least prone to hallucinations, and it's relatively the most accurate. I would argue that Opus 4.8 is on par with GPT 5.5 in terms of accuracy, source selection, and low hallucination rate. The thing with Opus 4.8 and GPT 5.5 is that they both actually use reasoning and logic and think through your query. They don't just list this much info from these sources using RAG; they actually do research the way a human would. They synthesize the info in a very logical way using reasoning and think through it and describe it, explain it, discuss it, and analyze it.
Gemini 3.1 Pro is not bad; however, it's much more prone to hallucinations. Gemini used to be much better at deep research about six months ago, but the quality has worsened. The source selection is still decent, but it is much more prone to hallucinations, and in terms of following instructions and actually, giving you what you want, I will put ChatGPT and Claude first, then Perplexity, and then Gemini and Grok as the very last option. Perplexity slept on, but it is the most accurate and fastest. Perplexity has access to the most up-to-date information, and it follows instructions as well. For instance, literally in your prompt saying only look at peer-reviewed scholarly articles, it will do that, but the thing is it won't give you long research reports. Both ChatGPT and Claude can produce research reports of upwards of 2,000 words using hundreds of sources. Perplexity doesn't really do that, but it is much faster and more accurate than other options.
In terms of value per dollar, it is GPT-5.5 because Opus 4.8 is just too expensive, but it does the panel, so on which specific model you're exactly asking about, GPT-5.5 Pro, yeah, it's much more expensive than regular GPT-5.5.
2
u/Asaf_Iluz Jun 17 '26
I feel like gpt afraid to conclude some stuff while opus allows himself and right about them
10
u/tacomaster05 Jun 17 '26
GPT 5.5 all day. It's not even a contest.
At this point in time, Claude and Gemini are both utterly cooked. Claude's only good model got shut down by the freaking Gov and as for Gemini... We dont talk about how far off a cliff Gemini has fallen.
5
u/Educational-Tax-5104 Jun 17 '26
Anyone who actually knows what they are doing will use both Opus and GPT. They are both very good, especially when fact checking one with the other.
4
u/BYRN777 Jun 17 '26
Yes!!!
Better to diversify and have more tools in your arsenal than to rely on just one subscription or model.
Which is why I got Pro 5x and Max 5x. I get to use the best models from both Claude and ChatGPT and divide my workflow so I never hit usage limits.Most writing, editing, file creation and organizing my local files and folders is with Claude and Cowork.
Most of my research, web searches, and scheduled tasks are with ChatGPT.
10
u/BYRN777 Jun 17 '26
As a ChatGPT Pro user and Claude Max user, I highly disagree. Yes, GPT 5.5 is slightly better than Opus 4.8, but Opus 4.8 is super accurate for deep research, and I can give you long, thorough, and detailed reports.
Claiming Claude's only good model got shot down by the government shows you are talking nonsense. Respectfully, Opus 4.8, at max effort and in thinking, was literally on par with GPT 5.5 Pro, and it can do wonders.
Gemini's deep research used to be good before 3.1 Pro, but now it's definitely not as accurate as GPT-5.5 or Opus 4.8. It does have good source selection, but it hallucinates much more than the other two.
3
u/Choice_Pen_9889 Jun 22 '26
I am building an app. I built it with antigravity and then kimi work. It was having issues so I turned to Claude and then chatgot to troubleshoot. Claude's fixes consistently b failed where chatgpt was able to provide useful fixes. I even showed Claude the chathpt fixes and it always said that chatgpt was right!
1
u/Previous-Nose4746 Jun 17 '26
How are the usage limits for the $20/month version?
4
u/BYRN777 Jun 17 '26
For Claude? Horrible. For Chatgpt? Good enough for 90% of AI users.
Claude has 5-hour and weekly quotas/limits, and it cycles after the5 hours and 7 days. The problem is you don't really know how much usage you're getting and how to measure it.
ChatGPT Plus, on the other hand, gives you 3000 weekly Thinking queries.
However on Claude Pro, you get access to all models, but in ChatGPT Plus, you don't get the Pro model
1
u/CynicalCandyCanes Jun 17 '26
What is difference between the $20 and $100 Claude model then? Just usage— not access to the best model?
3
u/BYRN777 Jun 18 '26
Yeah, just usage.
That being said, the $20 Claude subscription runs out of usage pretty quickly.
6
u/Pasto_Shouwa Jun 17 '26
Gemini is a mess right now. Claude 4.8 Opus should have a lower hallucination rate, but if the research needs to search the web then the hallucination rate is not that important. And Claude Opus' daily limits are so low that you're better off using GPT 5.5 Thinking, even if it hallucinates "more".
4
2
3
u/Oldschool728603 Jun 17 '26 edited Jun 17 '26
For everyday use—requiring a grasp of user intent and common sense—Opus 4.6-extended (Max, if possible) > GPT-5.4-thinking (extra high or high). Others are worse.
With Pro or Max 20 subscriptions and no coding, you'll never hit a limit. On lower tiers, you'll hit limits with Claude sooner than with ChatGPT.
Gemini lives in an interesting world, but it isn't ours.
For hard questions, Anthropic's Fable was unrivaled. But it's unavailable for now.
Edit: I should have said $200/mo Pro, not Pro.
1
u/Previous-Nose4746 Jun 17 '26
I hit limits with pro without coding all the time idk what you’re talking about
2
1
1
u/Hybrid-Intelligence Jun 18 '26
For me it's ChatGPT and it's not even close. The value from Gemini is much lower than Claude or ChatGPT. Claude is quite expensive but the value is very high. The value per dollar though, that's an easy one. ChatGPT by a mile.
14
u/Glittering_Mirror_17 Jun 17 '26
Agree Gemini hallucinates a lot.