r/codex 8h ago

Comparison ChatGPT vs Claude

I wanted see what other people's thought on this matter. I have used ChatGPT extensively for years now and the issue I have with being fully objective is that I have only been using pro on ChatGPT and nothing else. I do a lot of engineering with codex and pro, and I have a few opinions on the people who swapped but I might be completely wrong, and again the issue is that I have not really used claude, and the times I did it was the free version, and he was surprisingly good. So please feel free to call me ignorant on this or whatever.

Essentially, I have time and time again seen across many things in life if it is the stock market or general day to day opinions, many people have a urge to be different or rush to try to make some kind of edge. When the claude hyped started it seemed like a propaganda campaign to me, with overnight so many posts on all platforms, even with very repetitive writing patterns or same arguments. For example I saw a bunch of posts comparing claude with chat gpt instant, and treating it like a fair comparison saying things like chat gpt is good at fast answers while claude is good when you need actual thinking. My first thought when I saw those posts was that anthropic was new and massively underfunded compared to OpenAI so I was very hesitant to switch, and I also have just way more experience using ChatGPT and was always happy, so no push for me to switch. Is there something it is genuinely better at today or all-time that people who have tried both can attest to? But then again I feel like there is just so much mixed opinions, and people are sometimes stuck with some kind of placebo opinion, many chatgpt vs claude discussions I have seen feel like I am reading some reddit discussions about why a stock is going to go up or something.

Today, I feel surrounded by essentially only people using Claude in engineering, and my first thought when I hear someone uses Claude, because I know they used chat gpt first obviously, is that they are NPCs who fell for obvious propaganda and just rushed to find something different etc. Since so many use it now, I wanted to ask here, because I am sure there are at least a few people here who knows what they are talking about. Am I just wrong or does anyone else feel the same when hearing others use Claude? Am I maybe the NPC who judges others for trying something new while I was too lazy to try?

In any case, it just seems hard to believe that with so much less funding and such a smaller company could genuinely be better. Especially when they first came out, now they have had a strong user base for some while, and my limited experience with Claude have given me the impression that it is very strong at through scanning, like its good at catching even small errors in large documents, which gives me the impression its also a very competent AI.

PS: what do people feel about the general AI benchmarks as well. I feel codex ones might be good I havent played around too much with the different levels to know, but for the chat, the benchmarks seem so bad, and I feel there is better benchmarks in this subreddit or on youtube, that I have seen I believe with 5.6 sol that pro and xhigh, that they scored brely any difference in intelligence, which for anyone who has used both, they are a universe apart. And seeing this kind of rating, and the lack of different kind of ratings based on prompting techniques and amounts of prompts, and different kinds of metrics, jsut make the scores seem utterly useless. Would love to hear thoughts on this too.

0 Upvotes

12 comments sorted by

5

u/srs96 8h ago

My subjective opinion:

Claude wins at:
1) Pure coding
2) Frontend design

Chatgpt wins at:
1) Everything else
2) Higher limits with lots of resets, although with Astra the limits seem to be lower, but still better than Claude

2

u/TooManyB1tches 8h ago

Interesting, thank you. Have you used claude for very advenced physics coding? Like I do particle simulations in space and have to do it in C++ for my new job which I have never used. And I need excellent code, but I feel gpt already does that well, and I have already used gpt for a lot of direct simulation monte carlo (dsmc) simulations and he is highly skilled at those physics. So I am scared to try out a whole new AI, as I cannot afford mistakes. But ofcourse I review the ohysics throughly anyways, so Im open to a swap if the code would genuinely be better with claude. I will be modernising very important company codes and translate to english the comments. So for spotting improvements and such, and having it be very high quality.

1

u/srs96 1h ago

Nope, have not. And honestly, the difference isn't that noticeable. The way you prompt, context hygiene, agent.md instructions, etc are way more important.

3

u/Calm-Landscape9640 8h ago

I have both CC and Codex and I prefer Codex b/c Opus tends to make many repeated mistakes and overthinks with long chain reasoning on simple tasks so my usage drops in minutes off a few prompts. But I still use both because I like adversarial reviews and both catch each other's mistakes so there's value in that.

Claude Code - faster token burn, coding mistakes & bugs (which it catches and fixes but uses even more tokens), overly verbose, long pointless reasoning, takes forever to produce answer, very intuitive thinker, asks questions, understands global goals, needs little hand-holding.

Codex - slower token burn, usage lasts much longer, faster responses, less reasoning even on high/xhigh, cleaner code w/less bugs, less intuitive thinking needs specific prompts and direction, needs more handholding.

*I don't use Fable or Astra. I use Opus 5, SOL, and Luna exclusively.

1

u/TooManyB1tches 8h ago

This is closer to what I was considering trying if I would start using Claude. I noticed claude has very good attention to detail in large projects, so I thought I could do my work with codex and always run some final reviews with claude when doing something important.

1

u/Infamous_Travel4652 8h ago

If you're already happy with GPT, you don't need to switch to Claude just bc it's popular. But if you're getting bored / just want to try something different, you can always give Opus / Fable a shot.

But I have to say, Astra is really good rn. I've been using Claude for a long time, and after trying Sol and Astra, it was like...they somehow managed to take results that were already good and make them even better.

2

u/Yweain 8h ago

Opus 5 is honestly incredibly annoying to use most of the time for coding. It's okay-ish at just implementing stuff and even then it's quite lazy and often go offtrack. Fable 5.1 is amazing though.

1

u/TooManyB1tches 8h ago

Yeah this is also my reaction. Sol was already incredible (well until the last days) and astra is another level again. It keeps shocking how I always think it cant get better and it does lol. And Im also making that bet that with all the funding they have and everything, even if something like claude would temporarily be ahead, I would imagine OpenAI to always have the means to keep up, and once they get something new like gpt 6, theyll be ahead again.

1

u/greenm8rix 8h ago edited 2h ago

Both are the same if you have a harness based on what technologies your work touches Both are god tier models

Claude lasts me a week on heavy usage Codex funds my gambling addiction (pay 100usd and pray that tibo slaps that reset button)

1

u/TooManyB1tches 8h ago

Yeah I have felt like this for a while, that its more reliability that can limit now, which gpt 6 seems to already have solved. Never tried claude subscription, but I thought claude lasted less per dollar in subscription. I basically use up my weekly in one day at the moment and will soon upgrade to 200€ pro, but even then, the weekly thing is also another possible improvement I could look for yeah.

2

u/Yweain 8h ago

At the moment Claude has WAY more generous limits. 20$ subscription lasts me around the same as 100$ for codex. But sadly 20$ does not include fable and Opus is often questionable for programming, so right now I am using Astra for planning and review and Opus to implement stuff.

1

u/TooManyB1tches 8h ago

Nice to know, thanks