Baseten claims fastest inference for DeepSeek V4 Pro
baseten · x · 2026-08-19
Baseten claims to be the fastest inference provider for DeepSeek V4 Pro 0813 on Artificial Analysis, achieving 147 TPS. The company's engineers continue to optimize the model to deliver the highest throughput and lowest latency.
More from Infra
- Data center copper use to jump 75% to 1.3M tons by 2028 — Beth_Kindig · 2026-08-19
- Survey: What locked-in pricing and limits would make you switch LLM providers? — Resident-Pen-3757 · 2026-08-19
- Budget LLM provider OpenCodeGo slashes request limits by 400% — krrish253 · 2026-08-19
- Muon Optimizer Trains nanoGPT in Just 1.23 Minutes — dianarycai · 2026-08-19
- Cloudflare launches Monetization Gateway to charge AI agents per request via x402 — kleffew94 · 2026-08-19
- Open-source Profile v2.2 tunes vLLM from 81 to 421 tok/s on a single RTX 5090 — Inevitable-Diet-1870 · 2026-08-19