EN Submit a tool

DeepSeek V4 Flash

Rank #2 deepseek

deepseek/deepseek-v4-flash-20260423

Usage data date 2026-07-26 · updated daily

Usage and specs

Daily rank#2
Daily tokens6.4T
Daily requests611.6M
Daily change +18.3%
Context window1M
Released

What is DeepSeek V4 Flash

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and high-throughput workloads, while maintaining strong reasoning and coding performance. The model includes hybrid attention for efficient long-context processing. Reasoning efforts `high` and `xhigh` are supported; `xhigh` maps to max reasoning. It is well suited for applications such as coding assistants, chat systems, and agent workflows where responsiveness and cost efficiency are important.

DeepSeek V4 Flash capabilities

Input: Text   Output: Text

Reasoning output Reasoning Reasoning effort JSON output Structured outputs Tool choice Tool use

DeepSeek V4 Flash API pricing

Input · per 1M tokens$0.14
Output · per 1M tokens$0.28

List prices in USD; see each provider’s site for current terms.