DeepSeek V4 Pro
DeepSeek V4 Pro is a text generation model from DeepSeek with a 1,000,000-token context window and adjustable reasoning effort.
Long-context reasoning from DeepSeek's flagship series
DeepSeek V4 Pro is a large language model developed by DeepSeek and released in April 2026. It is the latest flagship model in the DeepSeek V series, designed for text generation tasks with a context window of up to one million tokens and a maximum response size of 384,000 tokens. The model supports a configurable reasoning effort setting, allowing users to tune how much compute the model applies to a given problem.
DeepSeek V4 Pro is served through DeepInfra and is suited for tasks that require processing or generating large volumes of text, including long-document analysis, multi-step reasoning, and complex instruction following. Its sampling controls — including Top P, Top K, Min P, presence penalty, frequency penalty, repetition penalty, and seed — give developers fine-grained control over output behavior. The model carries the FLAGSHIP, LATEST, and REASONING tags, indicating it represents DeepSeek's current top-tier text generation offering.
What DeepSeek V4 Pro supports
Extended Context Window
Supports up to 1,000,000 tokens of context, enabling processing of very long documents or multi-turn conversations without truncation.
Configurable Reasoning Effort
Exposes a reasoning effort selector that lets users adjust how much computational effort the model applies when working through a problem.
Large Response Output
Supports a maximum response size of 384,000 tokens, allowing generation of lengthy documents, reports, or code files in a single call.
Sampling Parameter Control
Provides Top P, Top K, Min P, presence penalty, frequency penalty, and repetition penalty inputs for precise control over output distribution.
Reproducible Outputs
Accepts a seed parameter so that identical inputs with the same seed produce consistent, reproducible outputs across runs.
Multi-Step Reasoning
Tagged as a REASONING model, DeepSeek V4 Pro is designed to handle tasks requiring chained logical steps, such as math problems and complex instruction following.
Ready to build with DeepSeek V4 Pro?
Get Started FreeCommon questions about DeepSeek V4 Pro
What is the context window size for DeepSeek V4 Pro?
DeepSeek V4 Pro supports a context window of 1,000,000 tokens, which allows it to process very long documents or extended conversations in a single request.
What is the maximum response length this model can generate?
The model has a maximum response size of 384,000 tokens, making it suitable for generating lengthy outputs such as detailed reports or large code files.
What does the Reasoning Effort setting do?
The Reasoning Effort input is a selectable parameter that controls how much computational effort the model applies when processing a prompt, allowing users to balance response quality against speed.
Who publishes DeepSeek V4 Pro and when was it released?
DeepSeek V4 Pro is published by DeepSeek and was released in April 2026. It is served through DeepInfra on MindStudio.
Is pricing information available for DeepSeek V4 Pro?
Published pricing for DeepSeek V4 Pro is not listed in the current metadata. You should check MindStudio or DeepInfra directly for current pricing details.
Documentation & links
Parameters & options
Non-think for fast responses, High for complex problem-solving, Max to push reasoning to its fullest extent.
Nucleus sampling. Considers only tokens whose cumulative probability exceeds this threshold.
Limits sampling to the K most likely tokens at each step. Set to 0 to disable.
Minimum probability threshold relative to the most likely token.
Penalizes tokens that have already appeared in the output, encouraging new topics.
Penalizes tokens based on how often they have already appeared.
Penalizes repeated tokens. Values above 1 discourage repetition.
Explore similar models
Start building with DeepSeek V4 Pro
No API keys required. Create AI-powered workflows with DeepSeek V4 Pro in minutes — free.