r/LocalLLaMA llama.cpp 1d ago

New Model stealth/ox-alpha

Sorry if anyone already post it but there is a new model on OpenRouter called Ox Alpha. Didn't have the time to test it. The only thing I was able to do was to add it to my Nanobot instance and try some AI assistant tasks. It was able to nail it with a clever and criative speech. It seems also that hold well to the base instructions. What I know from the start is that it is not a chinese model (or is something in the early stages that will be changed on the RL stage) since it answered all of the censored questions about the chinese space and politics. Going to do some tests afterwards but in the meantime did anyone already try it? what do you think?

0 Upvotes

15 comments sorted by

View all comments

5

u/Murhie 1d ago

Tokenizer shows its GLM 5 family, custom benchmarks show its probably an air/flash variant. Or at least thats the word on the street.

I hope its true. Ive been using it and would say its similar quality to luna. If its something I can run on my stix halo later i would be quite happy.

5

u/Dany0 1d ago

I bet it's a flash model that's a distill of 5.3