Model Library

Explore Models

Production-ready models optimized for performance and cost. Deploy in seconds via our unified API.

We introduce DeepSeek-V4.1-Flash, a multimodal Mixture-of-Experts (MoE) model with 552B backbone parameters and support for contexts of up to one million tokens. The model natively processes images and text, and generates text autoregressively.

deepseek-ai/DeepSeek-V4.1-Flash
$0.300per 1M tokens

DeepSeek-V4-Flash with 284B parameters (13B activated) — both supporting a context length of one million tokens.

deepseek-ai/DeepSeek-V4-Flash
$0.140per 1M tokens

DeepSeek-V4 series, including two strong Mixture-of-Experts (MoE) language models — DeepSeek-V4-Pro with 1.6T parameters (49B activated) and DeepSeek-V4-Flash with 284B parameters (13B activated) — both supporting a context length of one million tokens.

deepseek-ai/DeepSeek-V4-Pro
$1.480per 1M tokens