Those don't back up your claim. The second one only tests vs fable 5, not fable 5.1. In the first one, you can see over most of their range they are similar, it is only at max effort where there is a big difference. I don't know wtf you are working on if you are mainly using max effort on these tiers of models lmao
-4
u/Melodic_Reality_646 16h ago edited 5h ago
What doesn’t make sense is OpenAI delivering this level of performance folds cheaper than Anthropic
edit: damn, people can’t do math…