1-Bit Quantized Version of Inkling Released

danielhanchen · x · 2026-07-16

The author announced a collaboration with Thinking Machines to create a Dynamic 1-bit **GGUF** quantized version of **Inkling**, highlighting the following: - A **86% reduction in size**, shrinking from **1.9TB** down to **270GB**. - Retains **74.2%** of the top-1% accuracy. - Added **vision and audio support**. - Provides GGUF files and related guides, emphasizing that it can run on hardware with around **280GB** of resources.

Original post →

More from Infra

Infra channel →