MODELS
by DeepSeek
DeepSeek-V4-Pro is a DeepSeek Mixture-of-Experts language model with 1.6T total parameters, 49B activated parameters, and a 1M-token context window.
DeepSeek-V4-Pro is part of the DeepSeek V4 Preview release. Supplied DeepSeek sources describe it as a text-modality Mixture-of-Experts model with 1.6T total parameters, 49B activated parameters, and a 1M context window. DeepSeek API documentation lists deepseek-v4-pro as an allowed chat-completion model value and documents API features including thinking mode, reasoning_effort high/max, JSON output, streaming, and function tool calls. DeepSeek’s model card names DeepSeek-V4-Pro and DeepSeek-V4-Flash, gives a release date of April 24, 2026, and states MIT-licensed weights with API terms for hosted use.
Use for evaluating a DeepSeek long-context model where the supplied sources list a 1M-token context window. Use for coding reasoning or tool-using chat-completion workflows that need DeepSeek API support for function tool calls streaming and JSON output. Use when comparing large MoE models by total parameters activated parameters context length and availability of downloadable weights.
Last updated Jun 1, 2026