DeepSeek V4-vision-exp launches API with ultra-fast speed and low cost
teortaxesTex · x · 2026-08-21
DeepSeek has released the V4-vision-exp model via API. It processes images using just 117-384 tokens at V4-Flash price points. Early reports highlight its magically fast inference speed, positioning it as a strong contender in terms of speed and cost, potentially disrupting rivals like the Whale Inference Fleet.
Related event: DeepSeek Launches V4-vision-exp: Fast and Low-Cost Vision Model(2 posts)→
More from Models
- LLM German Output Cringed: Reads Like It Was Written by Olaf Scholz — DominiqueCAPaul · 2026-08-21
- Fastest NVFP4 quant of Qwen3.8 27B released, 50% faster than Q4 on compatible hardware — ionsago · 2026-08-21
- DeepSeek launches V4-Flash-Vision-Exp, multimodal agent performance nears Opus 4.8 — deepseek_ai · 2026-08-21
- Jie Tang on scaling history: FLOPs were intelligence, parameters were knowledge — cedric_chee · 2026-08-21
- Leaked Benchmark Suggests MIMO V3 Pro Performance Rivals Fable — teortaxesTex · 2026-08-21
- Leaked GLM-5.4 Reportedly Surpasses GPT Thanks to Rapid RL Iteration — kimmonismus · 2026-08-21