D

DeepSeek: deepseek-v4-pro

deepseek-v4-pro
1M context384K outReasoningToolsCacheStructured
Released Apr 24, 2026Updated Jun 18, 2026

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding, and long-horizon agent workflows, with strong performance across knowledge, math, and software engineering benchmarks. Built on the same architecture as DeepSeek V4 Flash, it introduces a hybrid attention system for efficient long-context processing. Reasoning efforts high and xhigh are supported; xhigh maps to max reasoning. It is well suited for complex workloads such as full-codebase analysis, multi-step automation, and large-scale information synthesis, where both capability and efficiency are critical

Mode chatTokenizer DeepSeekQuantization fp8Deprecation Feb 2028deepseek-ai/DeepSeek-V4-Pro

Pricing

Input price
$0.75/ 1M tokens$1.7457% off
Output price
$1.50/ 1M tokens$3.4857% off
Compatible endpoints openaiVendor DeepSeek

Uptime

Performance

Loading performance data...

Usage & Ranking

Loading usage...

Supported parameters

All providers = every upstream serving this model supports it. Some providers = depends on which upstream handles the request. Default = the value sent when you leave the parameter unset.

ParameterProvidersDefault
frequency_penaltyAll providers-
include_reasoningAll providers-
logit_biasSome providers-
logprobsSome providers-
max_completion_tokensSome providers-
max_tokensAll providers-
min_pSome providers-
presence_penaltyAll providers-
reasoningAll providers-
reasoning_effortAll providers-
repetition_penaltySome providers-
response_formatAll providers-
seedSome providers-
stopAll providers-
structured_outputsAll providers-
temperatureAll providers1
tool_choiceAll providers-
toolsAll providers-
top_kAll providers-
top_logprobsSome providers-
top_pAll providers1

Frequently asked questions

Similar models