Privacy-focused LLM service Venice hits 250B daily tokens, up 2.5x in months
0xAllen_ · x · 2026-09-17
Erik Voorhees announced that Venice, a privacy-first LLM inference service, now processes 250 billion tokens daily, up from the 100 billion figure he cited earlier — a 2.5x jump in a short span, signaling rapid growth in uncensored/decentralized inference usage.
More from Infra
- Boson AI CEO Alexander Smola on Voice Agents, Avatars, Latency and Emotional Intelligence — TWIML AI Podcast · 2026-09-17
- DDRop attack drops DDR5 writes to break Intel TDX and AMD SEV-SNP confidential VMs — matthew_d_green · 2026-09-17
- PufferLib 5.0 hits 60M steps/sec single-GPU RL training, solves Breakout in under a second — jsuarez · 2026-09-17
- Crusoe signs multiyear cloud deal with Perplexity to expand AI chip rentals — inductionheads · 2026-09-17
- SGLang's Delta Router Replay slashes sync stalls in Kimi K2 agentic RL training — hsu_byron · 2026-09-17
- SemiAnalysis: Rubin NVL72 delivers 7x perf-per-watt over GB300, far above Jensen's 3x claim — inductionheads · 2026-09-17