Quantization-Aware Healing turns a 4-bit model into one that beats its 16-bit original
On August 25, 2026, Multiverse Computing published Quantization-Aware Healing, a recipe that, applied to a GPT-OSS 120B compressed to 60B and quantized to MXFP4, beats its bfloat16 version on 7 of 9 benchmarks. The method distills from the original model, not from the compressed checkpoint.