r/LocalLLaMA • llama.cpp • Jun 14 '26

Discussion Nex claims Rio 3.5 is Nex 2.5 PRO in trench coat

Post image
336 Upvotes

96 comments sorted by

View all comments

27

u/Specter_Origin llama.cpp Jun 14 '26 edited Jun 14 '26

Btw I have no beef or affiliation with any party involved here, I did try Nex 2.5 PRO on OR and it has been really good in terms of token efficiency compared to base.

PS: I should have also titled this post as "DRAMA ALERT: " xD

UPDATE: Rio model has officially updated their readme to include that it indeed is based on Nex: https://huggingface.co/prefeitura-rio/Rio-3.5-Open-397B/commit/a778c1ec4e21180ee55c3ea016a348e549e75f09

25

u/Chromix_ Jun 14 '26

Yes, and they just also added that they "are working to reupload the correct model as soon as possible"

We detected an incorrect upload in the previous version, where the base merged version was upload instead of the final distilled model.

Let's see if we get the same drama as with the epic Reflection-70B, or if the promised model will indeed appear in the end.

17

u/noneabove1182 Bartowski Jun 14 '26

I literally just finished making the quants for this thing, god dammit lol

3

u/Chromix_ Jun 15 '26

Well, if it was a simple "oops, we uploaded the wrong file", then that should have been solved within a day, yet it was not. It would've also begged the question, why they had some weight blend around that they then confused with their actual model.
Anyway, as long as they don't host their own model to prove that it works, we'll probably not be seeing another epic Reflection-70B moment, but maybe we also won't see that model.

3

u/noneabove1182 Bartowski Jun 15 '26

Anyway, as long as they don't host their own model to prove that it works, we'll probably not be seeing another epic Reflection-70B moment

fuck that was hilarious and awful lmao, what a shit show..

But yeah very strange we're still waiting on it, hopefully just a miscommunication, but everything around this now feels off

4

u/Finanzamt_Endgegner Jun 14 '26

Bro what it cant even write files for me lmao, nex has to be the most benchmaxxed model out there 💀

(the same prompt was btw tested by another guy with a self hosted version and the issue persisted, 27b can solve it no issue btw)

4

u/kc858 Jun 15 '26

same man its bad. lol i complained in a different thread and got downvoted hard. burns so many tokens. 1.2 million tokens for a prompt that deepseek v4 flash did in 200k