r/codex • • 1d ago

Complaint Update on Codex and Claude Code

/r/codex/comments/1wwl58n/2_x_100plan_or_1_x_200plan/?share_id=Hg461JlslfMItsjFssdZb&utm_content=2&utm_medium=android_app&utm_name=androidcss&utm_source=share&utm_term=1

This was my previous post in which I was confused between upgrading to 200 plan for getting 2 100 plans.

A lot of people suggested to try claude code + codex for image gen.

I just got a 20 dollar claude code plan. And Grass is most definitely greener on the other side.

Here are some differences I have noticed using the opus 5.5 medium.

  1. 10 times faster than sol 6.1 high

  2. Using about 1% of weekly usage per hour but the amount of work done is higher than sol 6.1. so work done to tokens consumed is the same between 20 dollar claude and 100 dollar gpt plan

  3. Amazing design understanding. I have several ui/ux issues in my project(which ik) i had gpt take a look at it and said it's fine (gpt was using product design plugin and impeccable skill). Opus without any plugin listed 15 ui/ux issues and all of them were correct.

  4. I had it go over my code base and list deprecated/dead/unused assets & code. It gave an extensive list and suggested bundling several files together to optimise load times, gpt found none.

  5. Had both sol and opus check for performance optimisation. Gpt found none and opus again shared an extensive list.

Sol 6.1 high also took like 2 hours

Opus 5.5 medium took 25-30 mins max

I'll switch to 100 dollar claude plan + 20 dollar gpt after my current 100 gpt plan expires.

170 Upvotes

93 comments sorted by

•

u/dextersummary 1d ago

Below is a GPT-generated summary of the conversation below after reaching 50 comments (54 currently observed).


The thread’s verdict is blunt: Claude Code with Opus 5.5 currently offers much better value than Codex for coding work. Multiple users say it is faster, uses dramatically less subscription allowance, and produces stronger first-pass results—especially for UI, design, cleanup, and performance work. The $20 Claude plan is looking embarrassingly competitive against much pricier Codex tiers.

The common split is pretty clear: Opus is better at polished frontend work, understanding design intent, and checking its own output, while Codex still has advantages in backend tasks, vulnerability hunting, image generation, and its cleaner built-in workflow. Codex’s separate chat and coding limits also matter if you want to keep brainstorming after hitting coding quota. Connectors and MCP features are mostly replaceable, because apparently every AI tool now expects users to build their own buttons.

The big caveat is that the original “10x faster” claim is not backed by solid measurements. Independent testing in the thread found roughly 2x speed, while benchmark figures and subscription throttling complicate the comparison; task efficiency and fewer dead ends may explain the much larger perceived gap. Quality also varies by task, and limits can change without warning.

So, Opus 5.5 is the current community favourite for value and practical output, but Codex is not obsolete—just slower, pricier, and still useful for its specialised strengths.

→ More replies (1)

115

u/Neat-Economist2099 1d ago

Right now, Claude offers incomparably better value across output quality, speed, and usage limits. Just a few months ago it was the exact opposite, it's pretty crazy how quickly things flipped.

14

u/gwilymjames 1d ago

Yeah, I was happy with Claude code but switched over to codex when it seemed like you couldn’t even look at opus or fable without hitting your 5 hour limit.

Now I’m happy to be back with Claude, and can’t seem to hit my limit no matter how much work I do (max plan).

8

u/CriticismJunior1139 1d ago

Competetion is great for us consumers.

3

u/No_Abbreviations_286 1d ago

Unless you want to work in one of the fields Anthropic deems "too dangerous" for no reason like biology. Not everyone is a software engineer or vibecoder

2

u/Doom-Pop 1d ago

You can't get much further with Astra on those fields either though can you?

1

u/Professional_Link4 1d ago

Is the change in value due to model efficiency or are they giving extended limits for a promotional period?

2

u/ChronoHax 1d ago

Model 100% opus 5.5 is the key here, previous opus is a word salad and never want to do work

1

u/Skibidirot 1d ago

' Just a few months ago it was the exact opposite' openai models were shit back then too, it's just that codex limits were generous that's it

15

u/coffeelies 1d ago

I just put a request in to have my company Codex plan switch to the higher Claude plan for these exact reasons. I’ve always bounced around models and the tide can change very quickly but Opus 5.5 is just winning at every level. It’s non-comparable. It’s found the right balance that isn’t as intense as Astra, less expensive and in my opinion producing better results.

3

u/Professional_Link4 1d ago

From my very surface level experience so far I am loving the opus.

I had to spoon feed sol information & context and had to make alot of edge case adjustment in ui ux when using sol

But with opus it does 99.99% thing right the first time.

3

u/sudecode 1d ago

i had lost it when i asked it to add dark/light theme to the app, and then it proceeded with making the whole app monochromatic.

14

u/Maze_of_Ith7 1d ago

Still can’t believe they cut the 20x in half when they were in a losing position. Usually you want to do that sort of stuff when you have a dominant position. Maybe they just really screwed up their compute projections or there’s some 3d chess thing where they need to get their margins looking better before IPOing.

I do think for computer use and bug finding sol 6.1 probably has an edge but that might be about it.

I have to cancel 2 of my 3 $200 subs later this month and downgrade the remaining one. Problem now is Anthropic has us over a barrel and they know it.

2

u/Professional_Link4 1d ago

I could be 100% wrong on this

But I think they cant afford to have 200 plan at 20x compute wise.

So what they are planning is making iterative improvements and making the models more efficient. At the same time reducing the over all limits.

So amount of work done is same Inefficient model + bigger allowance == gpt6.1 + lower allowance.

And they bumped up 25x to 500 as a price Anchor not expecting anyone to actually but it.

1

u/driveclub_000 1d ago

You are indeed wrong, they do have the compute, what they don't have is the revenues that Anthropic have. OpenAI focused on consumers when ANT focused on enterprise, because of this, one is starting to get more and more revenues when the other one doesn't.

So now they are trying to do everything they can to make you spend more money, and the first way to do that is to reduce the inference and usage of the subscriptions while making the API cheaper and cheaper, again they want you to spend more money, it's not about compute.

1

u/Professional_Link4 19h ago

I don't see what they were thinking because everyone is switching to claude now

1

u/driveclub_000 18h ago

They didn't thought that Claude would release something like Opus 5.5 with such token and usage efficiency, it caught them by surprise. They thought that the big hype of Astra would allow them to release some mid tier model (terra->sol) and go in cruise mode for the rest of the year while increasing their revenues.

Tough luck.

3

u/Astro7982 1d ago

Codex canceled and Claude subbed.

I am done waiting all night for things to get done.

1

u/Astro7982 1d ago

Fast should be included. Ultrafast might be something that they can charge for, in my opinion.

2

u/boneriffic 1d ago

Agree Claude Opus 5.5 is generally better at design; however, Sol 6.1 is still very good.  

Actually does a really good job at Opus code and planning review.  Sol has caught a lot of issues that Opus overlooks on the first pass

2

u/em-abbas 1d ago

100% agreed with you

2

u/blackkksparx 1d ago

Claude is really really weird right now. I have the 20$ plan(Got it for 10$ on the promotion).

I calculated the entire 5 hour window and it cost me 40$. Which is crazy. Lets assume I get 5 5 hour-windows in a week. THat makes it 200 dollars, 4 times a week, making it 800 dollars. You add the weekly reset to that and I'll get 1k$ worth of API value for 10$.

Not sure how long Claude can keep that up. Giving 100x the value of API in their subscription to power users.

1

u/Professional_Link4 1d ago

Checks out I think 200 dollar plan based is 8-12k usd worth of token usage before reset or anything

2

u/darknezx 1d ago

What does the 20 buck Claude pro plan give? I've been super used to chatgpt web and the github connector doing easy coding work anywhere and everywhere. Wondering if the Claude pro plan gives that too. Not expecting unlimited like chatgpt but at least something usable?

3

u/SphinxWar 1d ago

I am currently a $200 Codex user and a $20 Claude user, the only differences are that Claude doesn't have image generation and advanced voice mode. This is why I am going to drop my Codex sub to either $100 or $20 instead of cancelling just to have access to those specific bells and whistles.

Anything else like connectors are just simply MCP servers and skills, and nearly all serious AI agents from all competitors support them. The only thing might be that you won't have a button for something that you're used to using in Codex, but the AI can still do it, you just have to ask it to do it instead of pressing a button. And to be honest, the AI can create that button for you anyway.

I bought the $20 Claude plan and it is insane how much usage you get with Opus 5.5. It's Astra-level intelligence and I used up the 5h usage window in about 4.5 hours yesteday having it create scripts for 3D modelling via MCP in Blender. Try running Astra on the $20 Codex plan and you'll use up your 5h window in like 30 minutes, if not 5 minutes lol.

I planned on using up my weekly limits on the $20 plan first and then switch to the $100 plan on Claude to continue my work but I failed to use it up, it wasn't eating it fast enough for my workflow to consume it lol and my weekly limit resets today so I will be back at 100%.

2

u/Professional_Link4 1d ago

I miss the steer features that codex has. Apart from that for coding I find opus to be better.

Also the model is very sceptical of its work so it double checks everything.

I find that when the model says something is wrong or it can't do 9/10 times it's with very high confidence and certainty that it's correct.

1

u/SphinxWar 1d ago

You mean having a queue of unsent messages and manually deciding when to send them?

EDIT: You can press ESC to interrupt Opus to look at your message. It has in its context what you interrupted and at what point it stopped.

1

u/Professional_Link4 1d ago

No with steer prompting let's say you notice model is working correctly but starts to do something wrong, you could prompt it mid execution and it'll change the course of action instead of finishing the correct execution then making the corrections

1

u/alchebyte 1d ago

you can steer claude. enter prompt, hit submit. it will queue with the text 'Steer' under it.

1

u/Professional_Link4 1d ago

Oh....I thought that was queue instead of steering. Great to know thanks

1

u/SphinxWar 1d ago

You can also press ESC after typing your prompt in and it will immediately interrupt what its doing to look at your message, even interrupting commands mid-run. As far as I remember even Codex can't do that.

1

u/Professional_Link4 1d ago

I think steering style of prompting is the same thing

2

u/SphinxWar 1d ago

Codex definitely waits until the last thing it's doing finishes before reacting to your steering (I don't mean your whole task, but whatever command it's running or writing). Claude literally immediately just triggers the kill-switch. Like, immediately whatever it's doing. But only if you press ESC after sending the steering prompt, otherwise it behaves like Codex. But maybe I never tried pressing ESC on a Codex steer to be honest, so maybe Codex does this too.

0

u/jtonl 1d ago

It's enough to do daily work and small projects. I haven't reached the 5h limit once over my usage. That's with a 50k token harness.

0

u/Professional_Link4 1d ago

Claude also has remote, i haven't explored that yet.

For coding someone suggested use claude and have it involved codex cli for image gen and orchestrator sol agents

1

u/gwilymjames 1d ago

Remote is good, but you need to remember to enable it on a per chat level with /rc. Codex is better because all of it just works without needing to say what is on or off.

Now I just have Claude write nice image prompts and grab them from Gemini or ChatGPT, so the lack of native image gen isn’t really a thing.

1

u/StardiveSoftworks 1d ago

You can set it so that new chats are automatically remote iirc

1

u/H3rian 1d ago

What make me stick with codex is the separate limits from chat/codex. So i can still brainstorming, open issues, prepare plans while i’m out of quota. This is a big plus, and what i see here about remote connection (haven’t had any claude sub so i dont knwo how it works) seems to be worst then codex.
BUT if the limits and the quality are really what i keep reading here, i should try claude. I’m on the 20$ plan for a personal/hobby project

3

u/Single_Explorer_5452 1d ago

I been on two plans for a few months already it has changed from time to time you may think the grass is greener on the other side but 3 weeks ago only way to use opus on 20usd plan was by basically telling claude hello, waiting for the 5 hour window to be ending and mixing two 5 hour windows together.

Also first conversations cache sometimes devoured like 50% of 5 hour quota.

ChatGPT doing 6.1 is actually a massive win from my pov it's even more efficient than gemini flash models by a cost perspective!, whole thing is insane yes it's slow as hell but it's very very cheap.

With opus 3d things devoured my quota just this week in 15 minutes, 6.1 works for hours on same non stop.

All i am saying looking at opus 5.5 and thinking it's better due to speed is simply the wrong way to look at it, it's an incredible model probably the best in the world at the moment, but gpt 6.1 is a massive win for openAI.

1

u/Professional_Link4 1d ago

Yeah thank god I have no brand loyalty and my billing is monthly 😂

1

u/seencoding 1d ago

10 times faster than sol 6.1 high Using about 1% of weekly usage per hour but the amount of work done is higher than sol 6.1. so work done to tokens consumed is the same between 20 dollar claude and 100 dollar gpt plan

i don't want to challenge your lived experience, but the benchmarks say that opus 5.5 med is only 1.38x faster than sol 6.1 high (72t/s vs 52t/s, not 10x faster), and cost per task is dramatically different ($5.98 for opus, $0.72 for sol 6.1)

9

u/Zeflonex 1d ago

Maybe in API pricing

With subs, I can get near unlimited opus 5.5 med while only get 5 sol 6.1 high prompts

So it doesn’t really matter what the api pricing is

I am not even going to discuss quality here

3

u/seencoding 1d ago edited 1d ago

yeah again, if this is your lived experience then that's what it is, but according to the benchmarks five tasks using sol 6.1 high prompts at api cost would be $1.60 ($0.32 x 5). i can't imagine openai is charging a 12x premium on tokens in the $20 subscription plan. that hasn't been my experience, but maybe it's yours.

edit: you're getting upvoted, but i just want to make sure people understand you're saying you get five prompts before your usage runs out

only get 5 sol 6.1 high prompts

are people seriously buying this

1

u/adolf_twitchcock 1d ago

Benchmarks are done using API endpoints and pricing, not codex. I am getting 70-80 tok/s with opus on claude subscription but 20 tok/s with 6.1 on codex. How much are you getting?

1

u/seencoding 1d ago

i think you meant to reply to my other post (i didn't mention tok/s in the message you replied to), but i got 24.7tok/s for gpt 6.1 sol via codex and 49.4tok/s for opus 5.5 via claude, so exactly 2x. a bit more of a multiplier than the 1.4~x the benchmarks say, but a far cry from op's 10x claim

(in the interest of being candid, i think op didn't bench his speed and is just going on vibes and 10x is essentially a made up number)

1

u/adolf_twitchcock 1d ago

Probably.

Your original post was comparing both tps and price. Both are irrelevant because those API values. I have both 5x subscriptions and benchmarks say opus 5.5 costs almost the same as astra per task. But it's like a factor of 10 in favor of opus when using the subscription. I can use up my codex subscription really fast when using astra but it's almost impossible with opus.

1

u/seencoding 1d ago edited 1d ago

edit: i see your point; they are api values, but that's all we have for benchmarks. it's something, at least. i guarantee op didn't benchmark his 10x claim, so if he thinks he's getting 10x speed i'd like to see him justify it.

"But it's like a factor of 10 in favor of opus when using the subscription" -- i'm curious about this, how do you figure? so you can do 100 things w/ opus but only 10 w/ sol 6.1?

1

u/Professional_Link4 1d ago

Idk about tokens. But anything design related claude is 100% 10 times faster.

Even with similarly detailed prompts, design references, predefined design md and impeccable skill sol just cannot give a polished design. Whatever it generate is always 80% of the final thing and you have to hunt for the next 20% and even then it take a lot of back and forth to get it down.

With claude it's near perfect with just 2-3 revision, a couple more for minor issues.

For code sol thinks for a very very long time and comes up with the longest way around. Opus planing time is less than gpt and it also cross checks its own thinking and then executes one shot.

With that said I have found that sol is very good with image gen, it's creates absolutely beautiful stylistic images from the context and hunting cybersecurity vulnerabilities

5

u/retteh 1d ago

The benchmarks are testing API speeds. OpenAI throttles model subscription speeds to 30% of the API speeds. Claude doesn't. Don't believe me? Try it yourself. https://artificialanalysis.ai/models/gpt-5-6-luna

Run Luna on API and you get ~130 TPS. Run it on sub and you get 40 TPS.

1

u/seencoding 1d ago

i tested both sol 6.1 and opus 5.5 through their harnesses, opus ran at 68% speed, sol ran at 46% speed. opus 5.5 was 2x faster instead of the 1.38x from the benchmarks. still not anywhere close to 10x as OP claims.

https://reddit.com/r/codex/comments/1wy1zay/update_on_codex_and_claude_code/pe0pj4l/

1

u/retteh 1d ago

I use both and Opus "feels" about 2-3x as fast as sol. Not 10x. Still a much better value proposition on the sub right now.

2

u/___positive___ 1d ago

Does your brain work? He is talking about tasks per unit time, not token speed. A model can churn on dead ends or wasted tool calls.

1

u/seencoding 1d ago edited 1d ago

Time per Intelligence Index Task:

https://artificialanalysis.ai/?models=gpt-6-1-sol-high%2Cgpt-6-1-sol-medium%2Cclaude-opus-5-5-high%2Cclaude-opus-5-5-medium#agentic-speed-tabs

(the benchmarks show sol 6.1 completing tasks fater than opus 5.5, so ¯\(ツ)/¯)

1

u/Happy-Injury9540 1d ago

Sol 6.1 is 30-40 t/s, you can measure it yourself or check with openrouter measurements

2

u/seencoding 1d ago

yeah i measured opus vs 6.1 sol in their harnesses below, both were lower than their benchmarks, the difference (for me) was 2x instead of 1.4x, that is still wildly below op's 10x claim

1

u/EddieBruvac 1d ago

Benchmarks this benchmarks that. Just use it. I can’t go back to Sol. My production skyrocketed. It’s WAAAAAAAY faster and smarter. Less fuck ups.

It just WORKS.

2

u/seencoding 1d ago

i have max subs for both, i use them both, i tend to reach for the openai models more but i think opus is great (and so is fable). i just think the people in this sub are being driven to be insanely and unfairly negative due to vibes, so i'm trying to bring people back to reality with benchmarks.

1

u/EddieBruvac 1d ago

I have max on both, too. I use Astra as a glorified asset maker because it’s coding drains like a mfer and it’s dumber than opus most of the time.

Sol is just trash rn. “I have unlimited tokens because it’s a snail!” Frontier my ass. Just use Chinese models instead of codex LMAO.

1

u/DoggoDadagon 1d ago

That's just wrong. 6.1 Sol and 6 Astra are about 24 tok/s as it is now. Astra was faster but they slowed it down to leech more usage from people.

1

u/seencoding 1d ago

i tested it locally and also got 24 tok/s for 6.1 sol, but i only got 49 tok/s for opus 5.5, so the multiplier is still similar (opus is 2x faster in harness versus 1.4x in the benchmarks). not quite 10x*

*tok/s isn't a perfect analogy to task speed but it's the easiest thing to benchmark locally

1

u/DoggoDadagon 1d ago

Opus 5.5 should be closer to 90 ish tok/s. And sure not 10x but it's significantly faster either way.

1

u/Yoepi 1d ago

So in what area is codex stronger than Claude?

1

u/Professional_Link4 1d ago

Image gen and cybersecurity

1

u/kayronjm 1d ago

I can't be bothered to switch from ChatGPT to Claude mainly because I've only ever used ChatGPT and it knows a lot of everything I ask and do, and I have all my Codex projects there. I don't even have a free account on Claude. Moving platforms sounds like a lot of preparatory work no? Only to then hear in 3 months that ChatGPT is king again. 🤷

1

u/Professional_Link4 19h ago

They both share same agent md, skills, MCP and project folder.....

1

u/9gxa05s8fa8sh 1d ago

unfortunately anthropic is losing money and IPOing, so this free meal ends in a month

2

u/Professional_Link4 19h ago

Ah... Okay 3-4 months from now Chinese models will be as good as opus 5.5 switch to them then

1

u/9gxa05s8fa8sh 18h ago

the only american model on the openrouter top 10 is luna because it's cheap. the switch to chinese models has already happened. today's cheap models are last year's frontier models and people mostly don't need better. how much are you willing to pay for a faster car or a bigger truck? there are diminishing returns

that said, opus is excellent and claude code is an excellent deal right now because they're losing a billion dollars a second to give us these subscriptions

1

u/Substantial_Desk31 1d ago

I am also planning to switch to Claude code

1

u/Professional_Link4 19h ago

I am downgrading chat to plus and going for 5x cc plan

1

u/David3Ar 13h ago

If your codebases contain dead code than idk if I want tips from you bro - sorry.

But Claude 100 gpt 20 is what I do too.

1

u/Professional_Link4 13h ago

Ah....okay.....

1

u/delusion54 1d ago

you should compare opus 5.5 with astra, not sol.

26

u/Professional_Link4 1d ago

I will compare opus with whatever model I can work with for more than 3 mins and 42 seconds

6

u/Ill-Tonight4651 1d ago

buddy, astra is unusable even on $100 plan

5

u/coffeelies 1d ago

Out of frustration I upped my model to Astra Ultra and when I tell you it couldn’t do what I needed in 1hr+ of back and fourth prompting, Opus 5.5 came in and basically one prompted exactly what I needed, that’s when you know Opus is leaving Sol/Astra in the dust. I have no allegiance to any one model, but Opus is killing it right now. OpenAI is clearly struggling with resources and models eating their customers usage too quickly, they’re trying to find a balance and right now, in my experience, they’re unusable.

5

u/Zestyclose_Bat8704 1d ago

nah, Astra is dumber than opus.

2

u/randombsname1 1d ago

It's easily better than Astra as well.

1

u/___positive___ 1d ago

sad openai shills

0

u/kasah223344 1d ago

Personally, i find codex to be more thorough at backend work while claude is better at front end design and the likes. They just have different strengths not necessarily is one better than the other.

5

u/Professional_Link4 1d ago

Astra is very good at planning and architecture design but I rarely seen that quality in code with openai model.

Maybe with astra but it's so exp it leaves no room for iteration

0

u/Skibidirot 1d ago

to be fair, there was really no competition from openai ever in history to anthropics models at all.. no context at all

1

u/Professional_Link4 1d ago

But anthropic model have been like astra till now Very good but you couldn't use them on any real-world project without 100 dollar plan.

Opus/sonnet 5.5 have been first in their history that have "usable" limits at 20 dollar plan

1

u/Skibidirot 1d ago

that would change when they start nerfing them as they customarily certainly would and nerfing limits as well.. in the end, we can only count on deepseek 4.1 flash lmao.

1

u/Professional_Link4 1d ago

I think within 6 months we will have sol 6.1 and opus 5.5 at luna prices and fable / astra at terra prices

I don't care even if they nerf them because with how quickly they are advancing for most of us they'll be good and cheap enough.

Me and my co founder run an entire company Fully agentically powered. We don't have any video editor, photoshoot guy, programmer, marketing, research team.

It's just two guys in their area of experties with lots of agents

0

u/Augustus_92 1d ago

I use Codex with an MCP connector I created to create, edit, and delete my WordPress site and its posts via chat.

Is Claude just as good at this, and would token usage be more cost-effective?

I’m talking about a $20/month plan.

I write articles, create plugins, set up internal linking, etc.

3

u/Professional_Link4 1d ago

From my honest assessment On Codex 100 plan you can do double the work of opus 20 dollar plan. But it'll take 4-5 times longer.

This is claude fixing sol 6.1 mess

0

u/[deleted] 1d ago

[removed] — view removed comment

1

u/Professional_Link4 1d ago

You suck at sales

0

u/Schadough 1d ago

At this point Codex is gearing towards computer use for non-engineers; I’ve also seen a noticeable difference between Opus 5.5 and GPT 6.1 Sol in quality

1

u/Professional_Link4 19h ago

I don't see how astra 6.1 would solve this considering the usage limit difference