Relace cuts Deepseek v3 inference costs by nearly 50% on OpenRouter

stuffyokodraws · x · 2026-08-25

Relace AI claims to have reduced serving costs for Deepseek v3 Flash on OpenRouter by nearly 2x compared to competitors. The optimizations are built on high-speed inference capabilities, reaching 10k tok/s Fast Apply and 50k tok/s Compact models. After two weeks of effort, Relace now tops the OpenRouter charts for Deepseek and is soliciting partnerships with coding agent developers offering free tiers.

Related event: Relace Cuts DeepSeek v4 Flash Pricing to a Fraction of Official Rates(2 posts)→

Original post →

More from Venture

Venture channel →