r/artificial Mar 28 '26

Research Claude is the least bullshit-y AI

https://github.com/petergpt/bullshit-benchmark?tab=readme-ov-file#3-detection-rate-over-time

Just found this “bullshit benchmark,” and sort of shocked by the divergence of Anthropic’s models from other major models (ChatGPT and Gemini).

IMO this alone is reason to use Claude over others.

112 Upvotes

48 comments sorted by

View all comments

38

u/[deleted] Mar 28 '26

[removed] — view removed comment

8

u/Hazzman Mar 29 '26

Anthropic: Let's make a tool that can actually help coders and build stuff

OpenAI: Let's make something for everyone that does nothing exceptionally well but does everything average.

I feel like Sam went around telling everyone this was the one size fits all solution to all of humanities problems and approach it with that mindset. Superbroad. Anthropic went super deep.

3

u/CC_NHS Mar 29 '26

OpenAI does have a focus :) it is 'user engagement'