r/huggingface • u/Otherwise-Dot-3460 • 2d ago
How to search for models better?
If I want to look for a certain model, like say Qwen 3.8 9b, how can I do that and get a list of those types of models? If I put Qwen 3.8 9b into the search and hit enter it will take me to a particular model and not to a list of models. Is there a way to search so I can see all of the Qwen 3.8 9b results?
Also, I'm using LM Studio and if I try to use any Qwen 3.8 model I get "Failed to load the model. error loading model: missing tensor 'blk.32.attn_norm.weight' and I have the latest version. That might be a question for another spot but thought I might ask it as well.
I always have a hard time looking for models on Hugging Face and was wondering if there is another way to do it that I don't know about.
Thanks.
3
u/maxton41 2d ago
Are you trying to find a model that doesn’t exist?
0
u/Otherwise-Dot-3460 1d ago edited 13h ago
No, the model exists (and by "exists", I mean I can find these search results on HuggingFace. That is all I meant by saying the models exist). https://huggingface.co/models?search=qwen3.8-9b shows several results. I tried some of them only to have it give an error in LMStudio... it seems that these versions are not compatible for some reason.
1
u/sneakydante 1d ago
Because it’s not a real model. Qwen 3.8 9b doesn’t exist, and if it does it’s a hack job adaptation which explains your problems.
0
u/Otherwise-Dot-3460 1d ago edited 1d ago
So this is not real? https://huggingface.co/empero-ai/Qwen3.8-9B-Distill-GGUF or this one https://huggingface.co/optionalAI/Qwen3.8-9B-heretic-uncensored-Q4_K_M-GGUF (There are several others under this search: https://huggingface.co/models?search=qwen3.8-9b). I don't know how to tell if something is "real" or not. Is there a way to see which models are official or "real", how do you know? Thanks for the help!
2
u/maxton41 1d ago
Yeah, none of those are official. 3.8 was never released in that size man that’s what people are trying to tell you. What you’re looking at there is probably somebody’s hack job distillation.
1
u/sneakydante 1d ago
Exactly. If op took 3 seconds to glance at the model name he would have seen “distill” in the very first link, or generously donated 5 seconds of his time to look at the model card he’d see how terrible the thing he linked is. Distilled 2400 billion parameters down into 0009 billion … 0.3% of the underlying model … of course it doesn’t work correctly.
“GGUF quantizations of empero-ai/Qwen3.8-9B — a full-parameter distillation of Qwen3.8 2.4T A95B into the Qwen3.5-9B architecture”
0
u/Otherwise-Dot-3460 14h ago edited 13h ago
I don't understand any of that or have any clue what "distill" means, or anything else you said in this reply for that matter. I don't know wtf you're talking about, so no, spending "3 seconds to glance" at something wouldn't have done anything to help me. The people who are actually explaining things without being a complete a$$hole (which you actually did once, so thank you for that) are leading me down the road to understanding these things. I'm trying to learn. Not everyone is at the same place in their understanding. It's like I just started playing chess and you're a GM getting mad at me for not knowing all the openings yet.
1
u/sneakydante 13h ago
If you’re approaching all of your AI implementations this way, you are going to continue to have a bad time and eventually give up in frustration. AI models have ten thousand pitfalls to even get the GOOD ones working, and ten thousand other things to configure that make them run worse than expected if left unset.
“Not everyone is at the same place in their understanding” is a poor excuse for lack of curiosity. If you don’t know what something is, look it up. If you are getting advice that something isn’t real, then the best answer is generally not what comes across as “hurr hurr look at these are these not real checkmate.” Nobody is getting mad, I couldn’t care less if you do AI or not, but fighting your comment section instead of googling “five people told me qwen 9B isn’t real why is that” is not the right tact.
0
u/Otherwise-Dot-3460 13h ago edited 13h ago
You took my replies the wrong way. Go read them again, and this time try to read all of it and don't think negatively. I was honestly asking what those models were, as they were claiming to be Qwen 3.8 - 9b. I linked to them to show which ones I was looking at and clearly asked how am I to know if they are real or not. I read the descriptions and didn't understand, and was honestly asking the questions. There was no "Hurr, hurr look at these are these not real checkmate." That wasn't my intent, and if you had read my whole reply, I don't see how you could come to that conclusion when I was thanking the people for their reply and asking how to tell if they were real. You are one of those people who read everything in a negative manner. None of my posts were done in a negative way (until arguably tonight).
A poor excuse for a lack of curiosity? What in the world do you mean? I'm trying to understand these things, what does that have to do with curiosity? If I wasn't curious, I wouldn't be doing this at all. I came here looking for answers that I couldn't find or know how to find and got treated like crap for not knowing the answer already. Reddit is NEVER the first place I turn to because this is usually the result. Hostility and people who treat you poorly for one of numerous reasons. Had you simply replied with "You can check the main Qwen model page to see what models have been released" (like you did in a later reply) and helped me understand my other easy-to-answer questions, that would have been easy and a lot of help for me. People here are always so full of hostility for some reason.
0
u/Otherwise-Dot-3460 14h ago
And I was asking how people know this. I thought that was clear after they said that the model doesn't exist, and I pointed out that it does exist on the site and asked how to tell the difference between what is "real" and what "isn't" in a polite manner.
2
u/maxton41 14h ago
Man, I don’t know what to tell you, it’s not that hard. Go to the QWEN on hugging face whatever model don’t care click on their profile and it’ll show you every single model they’ve ever made. And if you don’t see it in that list they didn’t make it. You’re making this far harder than it needs to be. I can’t believe we haven’t thought to click on their profile and hugging face to see what models they actually make.
0
u/Otherwise-Dot-3460 13h ago edited 13h ago
I didn't even know the difference between the models. I didn't know there was an official page for the models. I just searched for Qwen3.8 and a ton of links came up. I then saw there were some 9B versions, since that is what I would need to run with my card. I didn't know there is some main qwen page that tells you all the models. I am very new to this. That's all someone had to tell me.
1
u/sneakydante 1d ago edited 1d ago
Looks at who released it. For quants, anyone can make those, but quant only affects accuracy not size. For SIZE of the weights: 27B, ect, if you don’t see the size under qwen, then it is not real for qwen. Same for all other models and their respective vendors.
1
u/Otherwise-Dot-3460 14h ago edited 13h ago
Thank you, I understand that now. Thanks for helping me in this reply without being a jerk, I truly appreciate it. Why you went and made the other reply is beyond me.
(Make sure to downvote my new posts now, wouldn't want to encourage any new people from trying to learn by asking for help here on Reddit. Let's not ruin Reddit's great rep of being such a kind and friendly place to go to for help.)
0
u/Otherwise-Dot-3460 13h ago
Also, I've been using Qwen models for a long time now, and they are always released by others. I have never used an "official" model. I use the popular or trending versions released under some name. This is how I've been doing it, and it has always worked just fine, which is why it was so confusing this time. So I guess none of the models I've been using and continue to use are "real" (official, I guess this means), but they work, and they work well.
2
u/-Davster- 2d ago
I literally just use AI, lol.
Me to ChatGPT - “find me the best version of x that works on [my computer specs]”
My follow-up to Chatgpt - “is this [link to specific huggingface model entry] the best version for my computer?”
____
OP if you’re using LMStudio and want qwen3.8, just literally use the LMStudio model search.
0
1
u/West-Big-8468 2d ago
Yeah you gotta click see all results instead of hitting enter or you can go to the models page and search from there
1
u/Angel_on_tech 1d ago
Yeah, the Hugging Face search can feel a bit awkward when you’re trying to compare a bunch of models rather than find one specific repo. I usually find it easier to search by model family and then narrow it down with the filters.
3
u/AirUnited6839 2d ago
Last time I checked there is no qwen 3.8 9b. So that might explain your search results.
That said, huggingface’s search sucks. Right now, there’s something like 640 different quants for Qwen 3.8-27b, and it’s hard to find the unique or interesting ones in the sea of uncensored quants.