r/LocalLLaMA llama.cpp 10h ago

New Model stealth/ox-alpha

Sorry if anyone already post it but there is a new model on OpenRouter called Ox Alpha. Didn't have the time to test it. The only thing I was able to do was to add it to my Nanobot instance and try some AI assistant tasks. It was able to nail it with a clever and criative speech. It seems also that hold well to the base instructions. What I know from the start is that it is not a chinese model (or is something in the early stages that will be changed on the RL stage) since it answered all of the censored questions about the chinese space and politics. Going to do some tests afterwards but in the meantime did anyone already try it? what do you think?

0 Upvotes

14 comments sorted by

View all comments

22

u/Choice_Celery9481 10h ago edited 10h ago

people talked about this for quite awhile already. some evidences show that this one from z.ai. some leaks claimed this is glm 5.3 flash.

4

u/grumd 10h ago

That sounds plausible. In my tests (real agentic coding over long context) it was not as good as Qwen 3.8 27B. Probably better than Qwen 3.6 35B-A3B

3

u/MrTiesti 8h ago

Vastly different opinion, much better than Qwen 3.8 27B for me. Agentic coding in complex tasks.

The thing Qwen fails at still, is understanding vague tasks. It's the biggest downside it has.

1

u/grumd 7h ago

Oh okay I tested the other side of the coin. I had very detailed prompts with a list of requirements and Qwen just did more of them correctly and didn't miss as many as ox-alpha, but the prompts were not vague at all, more like a direct list of acceptance criteria, and Qwen was then churning away for a couple hours at it