Weaviate 1.39 adds 4-bit Rotational Quantization, cutting memory with minimal recall loss
CShorten30 · x · 2026-09-18
Vector database Weaviate shipped version 1.39 with 4-bit Rotational Quantization, slashing vector memory usage while keeping recall largely intact.
- New SIMD kernels and improved prefetching bring faster imports and search across the RQ family
- The accompanying blog covers the SIMD work, how recall holds as datasets scale, and benchmarks vs TurboQuant
- 8-bit RQ remains the default in Weaviate Cloud and now imports 16% faster
A solid option when RAM and cost reduction are the priority.
More from Infra
- Google Open-Sources Agent Substrate on GKE: 10x Density, 1,000+ Dormant Agents per Host — blaizedsouza · 2026-09-18
- Redditor crams six V100 GPUs into a standard full-tower case for local LLM inference — Odd_Caterpillar_2994 · 2026-09-18
- Crusoe raises $3.9B at $30.9B valuation to build data centers and modular AI factories — TechCrunch AI · 2026-09-18
- A 2.5-hour first-principles primer on the semiconductor supply chain worth your time — blaizedsouza · 2026-09-18
- YC F26's Dreamscale Labs Moves Robot AI Inference to the Cloud — ycombinator · 2026-09-18
- Community fine-tunes an MTP head for Bonsai 2 27B, ~1.25x inference speedup — cephaloform · 2026-09-18