LFM2.5 Q4\_0 Checkpoints from Quantization-Aware Distillation
Researchers have released LFM2.5 Q4_0 checkpoints, which utilize quantization-aware distillation to improve model performance at lower bit-widths. This method allows smaller, compressed models to maintain higher accuracy levels by incorporating knowledge from larger teacher models during the training process.
Covered by 1 source
- HHugging Face Blog↗2d ago