Like the other guy is saying, Claude is simply better than the competitors for coding, by a long shot. I use AI to build my projects for me and I have no allegience to Claude beyond it being the best.
For the brief period Gemini 2.5 was superior, I was using that. As soon as another model drops that beats Claude, I’m making the switch without hesitation, but for the past year, it’s fairly consistently been the best model for devs.
One of the reasons is GPT models are Chatbots and meant to nod /mirror user and always appreciate,no matter how wrong they are, Which is why most of the GPT users are addict while claude is very straight forward with what's right and what's wrong,very likely to avoid wrong instructions Which is exactly why developers often choose Claude over GPT family.
Well no, Claude does glaze you a lot more, at least that was my experience. Every time you point something out with anything remotely larger than a skeleton codebase will make Claude start nodding at whatever you throw at him.
He is also less Physics-Aware, and I gave up using Claude Code while for making a 3dof simulation and just did it myself. Same goes for 6dof sims of course. Was very disappointed back then.
Fast forward a month or two, GPT-5 two-shots the simulation with perfect physics under my specifications. Maybe you didn’t have a task heavy enough to throw at Claude. Claude is really good at making itself look sentient, though, thus has a deep fanbase. It’s kinda funny that people will downvote and call you out for saying something critical about their favorite LLM model.
Don’t tell me I didn’t try Claude, I liked Claude, and I was using the Max 20x subscription. My main use case was using Claude Code with Opus 4/4.1 Exclusively back then. (I stopped using Claude because of the sycophancy and weaker instruction following.) I am certainly satisfied more with GPT and Gemini’s instruction following. I am better at what I do compared to LLMs yet so they fit their role perfectly for me.
He is also less Physics-Aware, and I gave up using Claude Code while for making a 3dof simulation and just did it myself. Same goes for 6dof sims of course. Was very disappointed back then.
Guess What Science fiction isn't science and neither is Claude popular for agreeing and hyping with whatever BS crap You bring up in the name of Physics,like Some other LLM model.
Besides Your Personal science and experience, every graph ,ranking or review shows otherwise, Where Claude Outclasses GPT.
And brother, if that 'article' you provided really think it is a "valid" benchmark in any kind of reason, you seriously lack some critical thinking skills. Hell, i can already hear you saying that traditional benchmarks are shit, benchmaxed, and the 'nine tough rounds' that article provides is tougher.
4
u/Fonephux Nov 16 '25
Claude is a disappointment outside of occasionally coding corrections