Research Compressed 4-bit LLM outperforms full precision with QAH Quantization-Aware Healing distills 4-bit compressed LLMs from the original teacher. The 4-bit student beats its bfloat16 source on 7 of 9 benchmarks. Lars Cornelissen · Aug 25, 2026