River AI launches Rust-based inference serving DeepSeek/GLM Flash 20% below official rates
ibab · x · 2026-10-12
River AI launched production inference for open-weight models DeepSeek v4.1 Flash and GLM-5.3 Flash, priced 20% below official provider rates. The company built its own training and inference stack from scratch in Rust for higher efficiency, and new signups get $25 in free credits.
More from Infra
- AI buildout to cost $10.3 trillion to finance through 2032, topping all prior US investment booms — KyeGomezB · 2026-10-12
- Running a 456GB model on 192GB VRAM: offloaded inference hits 60-125 tok/s with 1M context — HankYeomans · 2026-10-12
- VitalOps launches agentic inference optimization, 2.6x median speedup — abhijithneil · 2026-10-12
- Running Qwen3.8 Flash-Next locally on AMD 7900 XTX at 500k context, 105-160 tok/s — human_in_the_looop · 2026-10-12
- Siemens Brings Nvidia Omniverse into Digital Twin Composer to Pave the Way for Physical AI — RevLebaredian · 2026-10-12
- OpenRouter token traffic explodes from 2T to 379T/month, open-weight models at 75% — Beth_Kindig · 2026-10-12