DeltaTensors Stores Fine-Tunes as Weight Deltas: 953MB → 294MB With Minimal Quality Loss
cupheadgamer · reddit · 2026-09-22
Open-source DeltaTensors manages fine-tuned checkpoints like Git, storing only the weight delta from the base model instead of full copies. On a Qwen2.5-0.5B fine-tune, a 953MB model compressed to 294MB with perplexity moving only 19.11 → 19.22. It works post-training (no LoRA or workflow change needed), supports chained deltas for version history, Hugging Face Trainer checkpoints, and streaming compression so you never load two full models. Repo: AaravGaurdev/deltatensors.
More from Infra
- Local Qwen 27B agent logs into Amazon and buys paper autonomously in one run — fuzhongkai · 2026-09-22
- 60 Minutes: US golf courses use more than twice as much water as data centers — SumitGup · 2026-09-22
- ROCm vs Vulkan on R9700 + Strix Halo: ROCm still wins for DeepSeek, Vulkan closes in — Hrethric · 2026-09-22
- Alibaba unveils new AI chip it calls China's most powerful — AIFlow_ML · 2026-09-22
- What coding models and quants do you run locally on DGX Spark? — be566 · 2026-09-22
- Intel Ships Day-0 OpenVINO Support for Qwen-Image-2.1 — Alibaba_Qwen · 2026-09-22