Neutrino-8B Hits HF Trending with Sub-2-bit Ternary Quantization

FermionResearch · hf · 2026-07-31

The FermionResearch/Neutrino-8B model is trending on Hugging Face. Based on Qwen3-8B, it focuses on extreme model compression using trtcv4 and sub-2-bit ternary quantization techniques to significantly reduce memory footprint and inference costs.

Original post →

More from Models

Models channel →