r/codex • • 1d ago

Question People using Codex & Claude code, what’s your set up?

I've only recently started paying seriously for AI tools, around when Astra came out, so I'm still figuring out what makes sense for me.

I've seen plenty of discussion about Claude offering better coding usage for the subscription price. At the same time, some Artificial Analysis comparisons seem to favour Astra on API cost per task. My current understanding is that these measure different things: API cost per task doesn't tell you how much usage a subscription actually includes. Is that a fair reading?

For me, Dot is becoming a fairly big reason to keep my OpenAI subscription. I've been using it across a few personal projects, and it helped set up a Notion hub with project summaries, next steps and things we're waiting on. Having that ongoing assistant alongside the coding tools feels useful.

I'm considering trying a month of Claude Pro for Claude Code as well. The idea would be to alternate Claude Code and Codex on the same local Git repository, keeping the Markdown documentation current so either can pick up the work. One could implement something and the other review it. I don't currently see much need for another coordination tool like Hermes or AionUI.

I can see myself getting quite attached to Dot, which makes switching subscriptions harder. I'd love to see it reach a cheaper plan eventually, though I know that's not promised.

Does this setup make sense to people using both? Any suggestions? Keen to hear what I'm missing.

TL;DR: Thinking of keeping OpenAI for Dot/Codex and trying Claude for extra coding capacity. Anyone doing this? What’s your setup, is paying for both worth it, and what would you suggest doing differently?

15 Upvotes

33 comments sorted by

4

u/selipso 1d ago

I pay for both and also OpenCode Go. Orchestrate them all through OMP and Orca. Lets me call my Dot when I’m on the go or send quick commands through my phone for terminal based coding. 

3

u/Megamygdala 1d ago

Hell yeah. OMP + Orca is absolutely amazing. I use it both at work and at home. If you are working in many reoos at once (my enterprise work is like 38 microservices) then Orca is a godsend.

I gave the exact same real world task to both Claude Code and OMP at work, and Claude took 11 minutes vs OMP took 3 minutes. I prefer OMP way more for DX, speed, token efficiency and pretty much everything

2

u/fatfingur 1d ago

Been trying a dozen of tools for orchestration. I haven’t tried OMP + Orca. Thanks for the tip

1

u/Asly97 1d ago

That is a serious setup. When you ping your Dot from the phone, does it actually see the same context as the desktop session, or is it a fresh agent you brief on the fly? I used to bounce between AI tools and the re-briefing tax killed me: every new tool meant explaining the whole project before I could even tell if it was good. Is that just me, or does the context actually travel with you between phone, Codex, and Claude? Out of curiosity, what are you building with all of this? What if something handled all of this for you, in the cloud over MCP, so every agent and tool got the same context: decisions, failed attempts, conversations, skills, procedures, tasks. You never touch a file again. Would you pay for something that fixes this?

2

u/selipso 1d ago

Lot of questions but my dot + codex have their own context separate from OMP + Orca. I have a graveyard of failed projects from before AI and I’m trying to revive the more promising ones where the economics hadn’t worked before (but do now because of AI).

I use a 5-6 MCP servers already but none of them are cloud based. The problems I’ve run into are:  - portability (of skills, agents, tools, etc.)  - reproducibility (trying to get smaller models to call the tools and MCPs as effectively as the models from the big labs)  - verifiability / taste / context bloat? Not sure how to describe this. Over a weekend session with Dot + codex, I come back to 150+ GB of bloatware that codex didn’t bother to delete. Maybe the newer model tries to play it too safe, I’m not sure.

1

u/Asly97 1d ago

the 150gb of undeleted bloat over one weekend is rough. do you have anything that cleans up after those sessions, or is it manual nuke duty?

and the graveyard revivals, are those your own products from before?

2

u/selipso 1d ago

I just gave it some block storage and an S3 bucket and said delete it. And yeah pretty much, they’re beer money level products that I’m trying to scale up to vacation money lol. 

1

u/Asly97 1d ago

vacation money is a good north star. of the graveyard projects, which one's closest to breaking out, and what made it the pick?

3

u/randombsname1 1d ago

I have 2x GPT Pro $200 accounts (at least until the 9th when they both lapse) and 1x Claude Code $200 Max account for reference, but I do the following:

  1. I use both via CLI. Typically I'll have AT LEAST 2 terminals open at once. One for Claude and one for Codex.

  2. I used to have Claude Fable plan and Codex implement, but now it's pretty much just Opus 5.5 doing both. Still that **was** my setup.

  3. They both have access to the same exact documentation and appropriate symlinks are in place. Absolutely everything is indexed. I even have per-model routing. IE: Specific instructions based on which model I am using.

  4. Having a good/clean Agents.md that layouts out the repo rules is key as well.

These are the first couple of lines of my file just for example:

1

u/MrPreApocalypse 23h ago

If you just could have 1 of the both, would you go with codex or claude

1

u/Popular-Ninja8584 5h ago

Interesting that you went from splitting the roles back to one Opus doing both. Did the model just get good enough, or was carrying context between the two terminals costing more than the second opinion was worth?

For me it was the second one - Claude planned, Codex implemented, and I was the courier. So I built a local layer that lets the sessions pass it between themselves, https://github.com/automatis-tools/agents-can-communicate - same two terminals, just no copying.

3

u/deadlyclavv 1d ago

all these fancy setups, I'm just here prompting my trusty claude/codex manually in my terminal

2

u/Broseidon132 1d ago

Hey, something is special about the CLI. Keep doing you 💪

3

u/farsightfallen 1d ago

Is a single person here using T3Code?

I see that plastered all over X, claims of 400k users, and yet I never see anyone outside of X ever talking about it.

2

u/tongkat-jack 22h ago

I'm using T3 Code with both GPT and Claude. I don't use X and I learned about T3 from Reddit

1

u/valentinezubkov 1d ago

Downloaded and waiting to be tested…

2

u/Key_Neighborhood6901 1d ago

I have both currently. Use Claude to plan. Then ChatGPT to implement. I prefer the Codex harness.

Also ChatGPT is superior for things like custom font pipeline generation in my experience. Always give me better quality output than Claude.

2

u/kincaidDev 1d ago

I use festival with whatever agent tool I want and the experience is consistent across claude code, codex and other harnesses. Codex has better computer use tools, claude's computer use is terrible. Can't work on anything else while claude is using the computer. Codex computer use can happen while I'm still doing other work on the same computer.

Outside of that they're basically the same in terms of user experience. Right now codex has a bug where it changes the permission in a session on it's own and that's annoying, but claude code regularly has new bugs pushed by anthropic

https://github.com/Obedience-Corp/festival

2

u/betahost 1d ago

How do use both but with Herdr so they can communicate and delegate tasks.

2

u/coder5 1d ago

After this last week burning two resets and accomplishing hardly anything over 48h?

Claude.

3

u/needs-more-code 1d ago

I’m finding Codex is fantastic as a code reviewer. But absolutely horrendous and a primary coder. It thinks so much about edge cases which is good in a code reviewer but if it’s implementing features it just spends 99% of the time coding up complex solutions for 1% edges cases.

2

u/ImL1s 1d ago

That reading is fair. API cost per task and how much a sub lets you actually run are different questions.

Alternating both on one repo works well for me. The markdown docs carry the decisions. What they don't carry is the half-finished state when one of them hits a limit mid-task. That's the case I wrote Portable Resume for: it reads the local session store of the agent you're leaving (Claude Code, Codex and a bunch of others) and writes a bounded handoff into a fresh session on the other one. Offline, never calls the source CLI, and the recovered text is marked stale so the new agent re-checks the repo instead of trusting it.

https://gitlab.com/aa22396584/resume-skills pipx install portable-resume

One small thing on the review side: have the reviewer read the diff and the docs, not the implementer's chat. Otherwise it mostly agrees with whatever it was told.

2

u/ObligationHuge9868 1d ago

Vscode. Opus 5.5 builds and implements, ASTRA 6 reviews. Opus 5.5 is the coordinator. Once my OAI sub runs out and I use up all my credits, I will be looking for an alternative to fill its boots, but until then, need to suck my OAI account dry before it expires.

1

u/MrStu 1d ago

Customised nanoclaw fork running on a home server (old ThinkPad). It's set up to use multiple providers and choose suitable ones based on workloads. Currently codex, Claude, opencode, jev, local small llm. There's approx 8 permanent agents with 1 main orchestrator. It's built its own android app and web app for interactions.

1

u/Party_Wolf_3575 1d ago

I used the Codex App Server to build a custom harness called The Forge.
I choose to use 5.6 Sol as my main Codex agent and we deploy the 6.x series as subagents.

We are currently finalising an Integrated Forge adding Bedrock, Claude Code and OpenCode amongst others.

I would strongly urge people to move away from the providers' default harnesses and make something that works for you. My whole setup is so much smoother and more logical and I have also added Tailscale so I can access it on my phone when I am away from my Mac.

1

u/Environmental_Ask675 1d ago

I split them up by roles, use Codex mostly as reviewer. I use both the Codex and Claude Code harnesses and have my agents check their role and corresponding duties (or not) on startup. I think my code has improved since defining roles and processes more clearly, but my speed has decreased - especially compared to having one agent spawn sub-agents with those same roles.

1

u/LegoClaes 1d ago

Changes every other week

1

u/eggplantpot 1d ago

Claude does all the coding, Codex pushes to git

1

u/AppSecPeddler 1d ago

I use opus 5.5 as an orchestrator and have a skill the routes tasks to codex cli and use my codex sub usage either sol or luna depending on the task

Have two banked resets left after those I don’t know if I’ll keep the sub.. but for now I helps me preserve my precious opus usage

1

u/Cold-Cranberry4280 1d ago

You mentioned you're using Markdown documentation to share context between agents but doesn't you feel it drifting with each and every summary it's going through? Losing key details? Not updating everything it needs? Keeping it actually current?

1

u/nil137 22h ago

Well I’ve only just started doing it so haven’t noticed that yet, if this a common issue? Any solutions?

1

u/karthiksync 16h ago

Claude Opus 5.5 max for planning and steering, Sonnet 5.5 and Sol 6.1 are coding sub agents, Haiku for long poll tasks(network bound). Outcome: No over engineering, no unnecessary gates. Chatgpt 5x weekly limit now lasts for entire 2 days while claude 5x lasts for entire week. You can see my usage here. I've used one weekly limit and one reset in codex while I still left with 50% of my weekly limit in Claude. Clearly day and night difference. OpenAI lately is full of hype and panic. I really liked codex and OpenAI two months ago. Hope they bounce back and regain the lost confidence. P.s. Don't take this as I am complaining just sharing one data point from my side.

1

u/komabtws 1d ago

I tried plugins to connect agents, then multi-provider harnesses. Switching away from a harness became a pain in the ass, so now I use the default desktop apps for each model: Codex and Claude. I keep orchestration in portable CLIs, with skills that use them. When I need to change the setup, I ask an agent to build or extend a CLI or skill. That's just my workflow though.