DeepSeek V4 Flash
deepseek/deepseek-v4-flash-20260423
Usage data date 2026-07-26 · updated daily
Usage and specs
What is DeepSeek V4 Flash
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and high-throughput workloads, while maintaining strong reasoning and coding performance. The model includes hybrid attention for efficient long-context processing. Reasoning efforts `high` and `xhigh` are supported; `xhigh` maps to max reasoning. It is well suited for applications such as coding assistants, chat systems, and agent workflows where responsiveness and cost efficiency are important.
DeepSeek V4 Flash capabilities
Input: Text Output: Text
DeepSeek V4 Flash API pricing
List prices in USD; see each provider’s site for current terms.