Choose DeepSeek V4 Pro When
- A 44.3 Intelligence Index and 47.5 Coding Index in TokenHub’s max-effort profile
- $1.80 input and $3.50 output per million tokens
- A 1M context window and 384K maximum output for text-only workloads
DeepSeek V4 Pro from DeepSeek and DeepSeek V4 Flash from DeepSeek are shown side by side so you can compare pricing, model IDs, context and output limits, modalities, tool use, release details, and benchmark results where data is available.
Overall Recommendation
DeepSeek V4 Pro is the capability default, with higher TokenHub Intelligence and Coding indexes while retaining the same 1M context and 384K output. Choose Flash when roughly 12× lower input cost and 11.7× lower output cost matter more.
Recommendation based on DeepSeek V4 Pro (Reasoning, Max Effort) vs DeepSeek V4 Flash (Reasoning, Max Effort).
Name
Model id
Intro
Author
Released date
Context length
Max output tokens
Name
Model id
Intro
Author
Released date
Context length
Max output tokens
Name
Name
Model id
Model id
Intro
Intro
Author
Author
Released date
Released date
Context length
Context length
Max output tokens
Max output tokens
Input
Output
Cached input
Input
Output
Cached input
Input
Input
Output
Output
Cached input
Cached input
Reasoning
Knowledge
Attachment
Input modalities
Output modalities
Temperature
Tool use
Reasoning
Knowledge
Attachment
Input modalities
Output modalities
Temperature
Tool use
Reasoning
Reasoning
Knowledge
Knowledge
Attachment
Attachment
Input modalities
Input modalities
Output modalities
Output modalities
Temperature
Temperature
Tool use
Tool use
Intelligence
Coding
Intelligence
Coding
GPQA
HLE
GPQA
HLE
SciCode
Terminal-Bench Hard
SciCode
Terminal-Bench Hard
IFBench
AA-LCR
Tau2
IFBench
AA-LCR
Tau2
DeepSeek V4 Pro and DeepSeek V4 Flash should be compared by workload. This page places price, context, output limits, capabilities, and benchmark data side by side.
Use the pricing rows above to compare input, output, and cached input prices for DeepSeek V4 Pro and DeepSeek V4 Flash.
DeepSeek V4 Pro lists 1M context length, while DeepSeek V4 Flash lists 1M.
If both models are available in the TokenHub catalog, you can route requests to DeepSeek V4 Pro and DeepSeek V4 Flash through the TokenHub API.