DeepSeek releases V4.1-Flash: smallest model in new family with native vision, $0.15/M tokens

aziz4ai · x · 2026-09-18

DeepSeek officially introduced DeepSeek-V4.1-Flash, the smallest model in its new architecture family, featuring native visual understanding and designed for greater capability, faster inference and higher throughput as a stepping stone to larger models.

It's already live on inference partner Runware from $0.15 per 1M input tokens via an OpenAI-compatible chat completions endpoint.

Related event: DeepSeek Unveils V4.1-Flash, Smallest Model with Native Vision(2 posts)→

Original post →

More from Models

Models channel →