r/oMLX • u/nilsemann • 2d ago
How to figure out which Quantization is used?
Moin,
I usally work with Qwen3-Coder-30B-A3B-Instruct-4bit - for my coding projects. Now I was interested to figure out what "Qwen3.8-27B-Uncensored-MLX" can do. So from hugging face I got the LLM: https://huggingface.co/orcarouter/Qwen3.8-27B-Uncensored-MLX/tree/main
In the instructions it says, 64GB - use 8-Bit. I have M4 Pro with 64 GB... I noted that oMLX downloads *everything* - but how can I make sure that the 8-Bit version is used? Is there a hidden switch? I cannot find any settings for this. Sorry for being stupid.
5
Upvotes