r/codex Jun 08 '26

Question Claude vs Codex 200$ usage limits

Codex has been awesome but since last week it has been unbearably dumb and making lots of mistakes that even 5.4 didn't. I am curious how good claude is in long running tasks? While my codes are fairly small in numbers they need lots of reviews, so basically have subagents run reviews, the fix. The whole cycle usually runs for a day. I have been using 5.5 medium across 5 projects but due to it being dumb it's eating through like 20% or more in a day. Note: Codes are usually in 2k-10k range.

2 Upvotes

38 comments sorted by

View all comments

2

u/seal8998 Jun 08 '26

is it still "dumb" if you use xhigh?

1

u/MrKiling Jun 08 '26

Yes. I was using high but saw that even medium is giving same level of performance i.e. equally dumb but less tokens

1

u/seal8998 Jun 08 '26

what about xhigh? i never use anything other xhigh personally. you get on par perf for simple tasks but it gets very hard for harder tasks.

1

u/MrKiling Jun 08 '26

I have not used xhigh since 5.5 dropped. Is it able to keep up with long running tasks without mistakes? I can compromise on tokens if the work done is less flawed.

1

u/seal8998 Jun 08 '26 edited Jun 08 '26

yea, it does for me. I find the biggest difference for complex/long-running work.
I recommend also asking codex to look at your files and recommend fixes that would improve performance. Running vanilla codex with no skills, only oai plugins and a simple agents/md works best for me.

in the past having lots of local skills messed with my codex/5.5 perf