The LFM2.5 model now offers Q4_0 checkpoints created via quantization-aware distillation. This technique reduces precision while maintaining higher accuracy than standard post-training quantization. Developers can now deploy these weights on consumer hardware with minimal performance loss. It is an incremental optimization for edge deployment rather than a fundamental architectural shift.