engy.ai Launches Kimi K3 with Lowest Inference Pricing
markjeffrey · x · 2026-08-05
Decentralized inference platform engy.ai announced the launch of Kimi K3, claiming to offer the lowest usage costs globally.
- Platform: Built for the agentic economy, providing high cache-hit rates and low-cost inference for open-source models with zero data retention.
- Pricing (per 1M tokens):
- Kimi K3: Input $1.5, Output $7.5, Cached $0.15
- GLM-5.2: Input $0.68, Output $1.5
- Qwen3.6-35b: Input $0.045, Output $0.30
More from Infra
- Nvidia's Strategic Shift: From Chip Sales to Cloud Revenue Sharing — GavinSBaker · 2026-08-05
- llama.cpp PR Caches Hot MoE Experts on GPU, Doubling Inference Speed on 8GB VRAM — BTA_Labs · 2026-08-05
- Big Tech Locks in Over $1 Trillion in Future AI Data Center Lease Payments — Polymarket · 2026-08-05
- YC-Backed Lamb Labs Claims 63x Higher Efficiency for AI Inference Chips vs GPUs — ycombinator · 2026-08-05
- DSpark Open-Sources Speculative Decoding Path for Kimi K3 — ying11231 · 2026-08-05
- US Drafts Ban on Chinese Datacenter Components as Europe Pushes for Tech Sovereignty — nordicinst · 2026-08-05