River Inference launches serving DeepSeek V4.1 Flash and GLM 5.3 Flash 20% below official rates
Scobleizer · x · 2026-10-12
River AI has opened its inference service to the public, serving DeepSeek V4.1 Flash and GLM 5.3 Flash at rates 20% below official provider pricing.
More from Infra
- $2 ESP32 board runs Pi-hole-style DNS blocker with 140K domains in 0.7MB — LinusEkenstam · 2026-10-12
- Engineer joins NVIDIA's Groq LPU compilers team, working on multi-chip partitioning — blelbach · 2026-10-12
- Linus Ekenstam wants nothing less than a 100B-param model running on your phone — LinusEkenstam · 2026-10-12
- AI buildout to cost $10.3 trillion to finance through 2032, topping all prior US investment booms — KyeGomezB · 2026-10-12
- Fireworks: open models plus fine-tuning match closed ones — Cursor gets 13x faster inference — AI Engineer · 2026-10-12
- Running a 456GB model on 192GB VRAM: offloaded inference hits 60-125 tok/s with 1M context — HankYeomans · 2026-10-12