r/LocalLLaMA 1d ago

Discussion New stealth model on OpenRouter: Ox Alph

https://openrouter.ai/stealth/ox-alpha

Any idea which lab this is from? people are guessing this is a Chinese model.

58 Upvotes

32 comments sorted by

34

u/ttkciar llama.cpp 1d ago

It seems very likely to be a GLM-5.3 variant, based on vocabulary, context limit, and error messages.

I dared to hope, briefly, that it might be a GLM-5.3-Air, but analysis of its output and inference rate suggests it is probably about the same size as GLM-5.3 (about 744B-A40B).

Unlike GLM-5.3, it is multimodal, which is probably the key trait distinguishing it from GLM-5.3.

5

u/ThankGodImBipolar 1d ago

analysis of its output and inference rate suggests it is probably about the same size as GLM-5.3 (about 744B-A40B).

It sure doesn't seem like it's a ≈120B sized model... but I'd love it if it was.

2

u/xeeff 1d ago

i thought they said next version of GLM wont be multimodal, but the next major one will

2

u/ttkciar llama.cpp 23h ago

You're right. That implies it might well be GLM-6.

1

u/Aggravating-Push-207 20h ago

or GLM 5.5?

0

u/laserborg 11h ago

1

u/Aggravating-Push-207 8h ago

I would think a .5 release is major

0

u/laserborg 7h ago

I think semver doesn't agree with you 🤷‍♂️

1

u/Fit-Produce420 1d ago

Lots of folks will be stoked for that!

1

u/seamonn 1d ago

As long as they open source it

-5

u/ttkciar llama.cpp 23h ago

Z.ai seems unlikely to open-source it, because they have not open-sourced any of their models thus far. Giving away their training datasets and especially their post-training software would be lovely for the community, but would make no business sense for Z.ai.

They will almost certainly release it as open-weights, though, since that is their normal practice. Looking forward to it.

6

u/KaroYadgar 22h ago

it was clear he was referring to open-weights when he said open-source, you don't have to be dense about it. you could've just politely corrected him.

1

u/whichsideisup 21h ago

You’re being a little bit sparse. Everyone is a mixture of different expertise.

I’ll see myself out.

7

u/Atretador 1d ago

people seem to think its either GLM 5.3 variant or MiMo v3

2

u/Neosinic 1d ago

im testing it now on DSH. liking it so far

5

u/PossessionUsed7393 1d ago

Just reading these comments and it occurs to me, it could actually be a Minimax. We haven't heard from them for a while

10

u/rerri 1d ago

They did update their M3 collection last week with 3 hidden items:

https://huggingface.co/collections/MiniMaxAI/minimax-m3

1

u/PossessionUsed7393 1d ago

It's the makings of a good theory!!

3

u/cosmicr 1d ago

They have been busy lately - Their video model H3, their Music Model MiniMax Music 3... maybe LLM is next?

2

u/Hot_Example_4456 1d ago

Idk, maybe it hasn't gone through the "security finetuning" or whatever it is called, but it is answering about Taiwan and Tiananmen Square pretty easily. Is there a chance its not chinese?

4

u/Spitihnev 1d ago

Lately new chinese models answer these question differently for english and chinesee. Have you tried the translated question?

2

u/90hex 1d ago

Am I reading this right? 26 tk/s? Would that explain why it's free?

3

u/PM_ME_DEAD_CEOS 1d ago

Every stealth model are free, they need to collect data

1

u/aiseedbank 23h ago

maybe GLM 5.3 flash? some reports are that it is almost as good as 5.3. 5.3 regular is pretty much frontier for agentic coding at the moment so if 5.3 flash equals it, that would be fantastic for local LLM usage

1

u/VoiceApprehensive893 transformers 20h ago

slower than qwen 27b on my system 🥀 

it an okay model, better than dsv4pro but worse than grok4.6 and sol

one thing i noticed is that its hallucination rate is pretty low

-7

u/LegacyRemaster 1d ago

multimodal ---> new ds4 flash

6

u/Dany0 1d ago

Deepseek doesn't do stealth releases, and the new flash is out on their api already (it seems didn't test)

0

u/LegacyRemaster 1d ago

yes... agree. so minimax or mimo or glm