deepseek-v4-pro

deepseek first-party price GPQA 91%

Also listed as azure_ai/deepseek-v4-pro, dashscope/deepseek-v4-pro, deepinfra/deepseek-ai/DeepSeek-V4-Pro, deepseek/deepseek-v4-pro, fireworks_ai/accounts/fireworks/models/deepseek-v4-pro, fireworks_ai/deepseek-v4-pro, nebius/deepseek-ai/DeepSeek-V4-Pro, novita/deepseek/deepseek-v4-pro, openrouter/deepseek/deepseek-v4-pro

Input$1.32per 1M tokens
Output$3.96per 1M tokens
Cached input$0.044per 1M tokens
A 1,000-in / 300-out request costs $0.0025 First-party list price · deepseek · recorded 2026-08-20

Price over time

$0.500$1.00$2.002026-06-17today

The same capability, cheaper

deepseek-v4-pro scores 91% on GPQA Diamond. The cheapest model clearing the GPQA 90%+ bar today is glm-5-3-flash at $0.237 blended — 8.3× less than this model's $1.98. Blended = 75% input, 25% output. Capability is one benchmark, not the whole story.

91%
GPQA Diamond
78%
SWE-Bench Verified
97%
Mock AIME
1,000,000
Context window
2026-04-24
Released
DeepSeek
Developer
3
Price changes
2026-06-17
Tracked since
3.6× dearer
Since first record

Recorded price changes

One row per change, newest first. US dollars per million tokens. Each row links to the exact commit or archived page the figure was read from.

DateInputOutput Cached inBlendedChangeSource
2026-08-20$1.32$3.96$0.044$1.98↑ 1.5×litellm@10829ff
2026-07-01$1.74$0.200$1.35↑ 2.5×archived page
2026-06-17$0.435$0.870$0.0036$0.544litellm@1ccc1e5

How this was collected

Prices come from replaying the version history of LiteLLM's public pricing file and from vendor pricing pages archived by the Wayback Machine, then cleaned to remove metadata edits, same-day churn, unit slips and typo corrections. The full method and every script is in the repository. List pricing, not negotiated pricing.

See how the cost of a fixed level of capability has fallen →