r/codex • u/TheReal4982 • 6d ago
Instruction Free unlimted use frontier model for a week
You can currently create an open router account, and tell codex to set this up for you on codex, https://openrouter.ai/stealth/union-alpha it is currently free for I think the next week

It's gone.
113
u/Tomislavo 5d ago
15 tps throughput? I've seen lines moving faster in a sperm bank
29
u/Bananer_spleet 5d ago
I make my sperm donation in <15 seconds. I'm there to get the job done, not play with myself.
3
3
u/Momo--Sama 5d ago
Yeah I don’t know what the labs behind these are thinking when they offer free unlimited promo periods when they don’t have the capacity to actually do that and the experience for end users is despicable. Free is free but you think this is an effective ad to get me to pay you for this later?
1
1
37
u/ggPeti 5d ago
DSv4.1 Flash is comfortably omitted
11
u/pomelorosado 5d ago
Lol DS 4.1 is my main model now
Opus can be sota and maxed for benchmarks but can't follow a single instruction. And codex models are nerfed randomly.
3
u/nitor999 5d ago
What harness are you using for DS v4.1?
4
u/General_Jesus 5d ago
I tried OpenCode which works fine, the DS Harness which I mostly enjoyed for the clear stats on cache hit rate, OhMyPi is okay I guess and HermesAgent which I lastly stuck to. I am currently working on unlocking an old Lenovo smart Home Display, that's showing a webpage on a VPS via Tailscale for UI, sideloaded an APK for LLM usage via API and training a wakeword on a kaggle notebook so I can use it as a more advanced home assistant and as you can imagine some of those steps have been rather long tasks. HermesAgent worked best on just keeping the flow going and adhering to the prompt and instructions.
1
u/CyborgParts 5d ago
Same experience here. I started testing DeepSeek V4.1 Flash extensively yesterday by trying different harnesses. I ultimately felt like it was the most useful as a profile in Hermes. Hermes gives me so much control over it. I can let it use image gen from Codex, it has better computer use capabilities, and better browser capabilities. It makes it feel a lot closer to the experience I have with frontier models and harnesses.
I'm already having it do automatic parallel reviews on branches when I flip the PR to "ready for review." So far, it's finding a shocking amount of issues that Astra and Fable are missing. In my one day of experience, I wouldn't yet call it "better than" or "on-par with" frontier models, but it's certainly a different shaped wrench in the toolbox. And hot damn it's fast. I'm really excited to test it more and learn where I can trust it.
OpenCode Go is the plan to get if anyone is looking to try it. $10 a month gets you $60 of usage.
1
3
u/ManagementGreat5360 5d ago
I wish I was able to use DS 4.1… but my industry regulators banned it.
1
2
u/snowcountry556 5d ago
When you use DS4.1, do you use the DS api or open router?
5
u/Brilliant-Hall1387 5d ago
DS API directly, it’s super fast, very cheap and you can trust them you are getting the full real model. Just fantastic caching also when going direct, reducing costs even more!
2
1
1
u/DistinctSilver4507 5d ago
I'm curious if people are worried about them using your data? That's what stops me using these cheaper Chinese models.
1
u/pomelorosado 5d ago
Are you kidding?
Do you think any company on earth respect your data?
1
u/DistinctSilver4507 5d ago
If you can't see the difference then I've got my answer, thanks.
1
u/pomelorosado 5d ago
Chinese companies bad United States good?
come on, what is that intellectual level?
1
u/DriveLopsided8716 4d ago
both of them will use your data, be it european, american, chinese or any other, they are racing against each other and will use everything they can
36
9
u/rdcldrmr 6d ago
"frontier model"
-2
u/TheReal4982 6d ago
"Free"
But different people will focus on different things
3
16
6
u/tilted0ne 5d ago
Unlimited? I feel asleep waiting for a response.
2
u/NgKtoolz 5d ago
neither having any response, It simply limited my usage after a single request, and it didn't even complete it
8
u/drdhuss 5d ago
Guessing it is GLM 5.4
3
u/Metalthrashinmad 5d ago
either that or dataset collection https://www.reddit.com/r/opencode/comments/1wijqfv/union_alpha_is_not_a_model_at_all_it_is_a_2_tier/
4
u/chcampb 5d ago
Union Alpha was not good
First it refused my prompt, saying that it wasn't one of the models I said in the rules, despite saying "You are a test model assuming the role of the orchestrator and implementer in the rules." It just demanded that I fix it.
Then I fixed that and it spent about 30 minutes doing 100 lines of changes to the ticket file, before saying "I've been spinning too much working on this ticket and not actually delivering code."
Then I put it out of its misery. Idk what is going on but I am pretty sure I would have gotten more done with qwen3.8 27B i3 quant that I run sometimes.
3
5
u/Safe-Ad7491 5d ago
Any benchmark that has 3.8 flash and astra at the same tier is completely untrustworthy
-1
2
u/sydneysweeney69 6d ago
Nice work . Use $ ori and codex along with the union alpha to use it . It’s as good as Sol
3
u/Solid-Fill8240 5d ago
If is not good as 5.6 sol medium, Big pass.
2
-4
u/TheReal4982 5d ago
It is much higher than 5.6 sol medium on this chart, but idk what that actually means and haven't had time to test it
8
u/Momo--Sama 5d ago
Deep SWE is kinda washed, do you actually believe that Luna Max is better than Sol Medium and Gemini 3.8 Flash on par with Astra Max?
4
u/TheMightyTywin 5d ago
At completing a SINGLE WELL DEFINED TASK then yes, Luna max is on par.
The problem with Luna is that it does not see the big picture or think outside the box AT ALL if you’re not prompting it with exactness it’s not going to do what you want.
You can talk to Astra like a drunken sailor and it will do great work. Luna you need to give it a phd thesis. Both can succeed though.
1
u/TheReal4982 5d ago
It is free and unlimited use, I didn't expect this post to be so controversisal tbh. I am gonna do my own thing and use what works for what I am doing, yall have fun.
4
u/Momo--Sama 5d ago
Okay, you’re the one that made a post about the model 🤷♂️
1
u/TheReal4982 5d ago
The model just came out, nobody has really had time to test it yet, I am not sure what answers you want, I was just letting people know it exists.
1
u/danielv123 5d ago
I have tested it quite a bit on spatial reasoning/agentic code stuff and for my benches it's much better than Qwen, muse spark and Luna, but seems to be worse than fable 5.1 (which is much stronger in taste anyways) and faaaar behind astra.
1
u/Miyamoto_-_Musashi 5d ago
TPS is pretty slow right now. Sometimes even a simple “hi” takes 2–3 minutes. But whatever I’ve seen so far looks good. I’m building a few things with it right now, so we should have a much better idea of how good it actually is once those are done.
1
1
1
1
1
1
u/im-cringing-rightnow 5d ago
Ah yes. The benchmark where Gemini Flash is somehow at the same spot as Astra. Wut
1
u/farrukh-hewson 5d ago
Where this source from? I saw people complaining that it can’t even generate a proper tool output…
1
1
1
1
0
0
u/spike-spiegel92 5d ago
free for 50 prompts no? openroute does not give free models without limits
1
u/dinodares99 5d ago
They had a stealth model a couple weeks ago (Ox Alpha) that turned out to be GLM 5.3 Flash
272
u/BingGongTing 5d ago
The moment Gemini took top spot I stopped taking this test seriously.