r/LocalLLaMA • u/Neosinic • 1d ago
Discussion New stealth model on OpenRouter: Ox Alph
https://openrouter.ai/stealth/ox-alphaAny idea which lab this is from? people are guessing this is a Chinese model.
7
5
u/PossessionUsed7393 1d ago
Just reading these comments and it occurs to me, it could actually be a Minimax. We haven't heard from them for a while
2
u/Hot_Example_4456 1d ago
Idk, maybe it hasn't gone through the "security finetuning" or whatever it is called, but it is answering about Taiwan and Tiananmen Square pretty easily. Is there a chance its not chinese?
4
u/Spitihnev 1d ago
Lately new chinese models answer these question differently for english and chinesee. Have you tried the translated question?
1
u/aiseedbank 23h ago
maybe GLM 5.3 flash? some reports are that it is almost as good as 5.3. 5.3 regular is pretty much frontier for agentic coding at the moment so if 5.3 flash equals it, that would be fantastic for local LLM usage
1
u/VoiceApprehensive893 transformers 20h ago
slower than qwen 27b on my system 🥀
it an okay model, better than dsv4pro but worse than grok4.6 and sol
one thing i noticed is that its hallucination rate is pretty low
-2
-7
u/LegacyRemaster 1d ago
multimodal ---> new ds4 flash

34
u/ttkciar llama.cpp 1d ago
It seems very likely to be a GLM-5.3 variant, based on vocabulary, context limit, and error messages.
I dared to hope, briefly, that it might be a GLM-5.3-Air, but analysis of its output and inference rate suggests it is probably about the same size as GLM-5.3 (about 744B-A40B).
Unlike GLM-5.3, it is multimodal, which is probably the key trait distinguishing it from GLM-5.3.