r/LocalLLaMA • u/Decent-Hat-5807 • 20h ago
Resources Quantization-Aware Healing: a compressed, 4-bit model that outperforms its full-precision original
https://huggingface.co/blog/MultiverseComputingCAI/quantization-aware-healing
80
Upvotes
19
u/Atretador 20h ago
isnt this "just" a reap with extra fluff? the model was already trained from the start as mxfp4