Token Price History / fireworks_ai

deepseek-v4-flash-0731

fireworks_ai resold GPQA 91% 2.9× cheaper

Also listed as azure_ai/DeepSeek-V4-Flash-0731, azure_ai/deepseek-v4-flash-0731, dashscope/deepseek-v4-flash-0731, databricks/databricks-deepseek-v4-flash-0731, deepinfra/deepseek-ai/DeepSeek-V4-Flash-0731, fireworks_ai/accounts/fireworks/models/deepseek-v4-flash-0731, fireworks_ai/deepseek-v4-flash-0731, nebius/deepseek-ai/DeepSeek-V4-Flash-0731, novita/deepseek/deepseek-v4-flash-0731, openrouter/deepseek/deepseek-v4-flash-0731

Input$0.140per 1M tokens
Output$0.030per 1M tokens
Cached inputn/anot offered
A 1,000-in / 300-out request costs $0.0001 Reseller price · fireworks_ai · recorded 2026-08-05

Price over time

$0.100$0.2002026-06-19today
91%
GPQA Diamond
94%
Mock AIME
1,048,576
Context window
2026-07-31
Released
DeepSeek
Developer
2
Price changes
2026-06-19
Tracked since
2.9× cheaper
Since first record

Recorded price changes

One row per change, newest first. US dollars per million tokens. Each row links to the exact commit or archived page the figure was read from.

DateInputOutput Cached inBlendedChangeSource
2026-08-05$0.140$0.030$0.113↓ 2.9×archived page
2026-06-19$0.220$0.660$0.0070$0.330litellm@3ead9d1

How this was collected

Prices come from replaying the version history of LiteLLM's public pricing file and from vendor pricing pages archived by the Wayback Machine, then cleaned to remove metadata edits, same-day churn, unit slips and typo corrections. The full method and every script is in the repository. List pricing, not negotiated pricing.

See how the cost of a fixed level of capability has fallen →