Liquid releases QAD quantized models, recovering 97% accuracy at 4-bit
helloiamleonie · x · 2026-08-19
Liquid AI released new 4-bit checkpoints trained with Quantization-Aware Distillation (QAD) for the LFM2.5 series (230M to 2.6B).
Technical Details:
- QAD distills a high-precision teacher model into a quantized student model to recover accuracy lost during quantization.
- Maintains the low memory footprint and high decode throughput of the Q40 format.
- All four checkpoints reach roughly 97% of their BF16 averages.
Developers can replace existing Q40 GGUFs directly for improved performance.
More from Infra
- Distributed Locking & TLA+ Verification at Modal — tokenbender · 2026-08-19
- New tool cargo-bsize helps analyze and reduce Rust binary size — charliermarsh · 2026-08-19
- Open MAX and Form Open Alliance to Unify NVIDIA, AMD, Trainium, Google TPU, and More — clattner_llvm · 2026-08-19
- Higgsfield Chooses Together AI for Inference on Dedicated Containers — togethercompute · 2026-08-19
- ModCon '26 wraps: Mojo is now open source and heterogeneous compute has a real software platform — clattner_llvm · 2026-08-19
- Grok 4.6 lands on Amazon Bedrock with 500k context and configurable reasoning — SpaceXAI · 2026-08-19