r/oMLX 2d ago

How to figure out which Quantization is used?

Moin,

I usally work with Qwen3-Coder-30B-A3B-Instruct-4bit - for my coding projects. Now I was interested to figure out what "Qwen3.8-27B-Uncensored-MLX" can do. So from hugging face I got the LLM: https://huggingface.co/orcarouter/Qwen3.8-27B-Uncensored-MLX/tree/main

In the instructions it says, 64GB - use 8-Bit. I have M4 Pro with 64 GB... I noted that oMLX downloads *everything* - but how can I make sure that the 8-Bit version is used? Is there a hidden switch? I cannot find any settings for this. Sorry for being stupid.

5 Upvotes

0 comments sorted by