gemini-1-5-pro-002

gemini first-party price GPQA 57% 2.4× cheaper

Also listed as gemini-1.5-pro-002, gemini/gemini-1.5-pro-002

Input$1.25per 1M tokens
Output$5.00per 1M tokens
Cached inputn/anot offered
A 1,000-in / 300-out request costs $0.0027 First-party list price · gemini · recorded 2024-11-07

Price over time

$2.00$5.002024-09-25today

The same capability, cheaper

gemini-1-5-pro-002 scores 57% on GPQA Diamond. The cheapest model clearing the GPQA 50%+ bar today is qwen3-7-flash at $0.055 blended — 39.8× less than this model's $2.19. Blended = 75% input, 25% output. Capability is one benchmark, not the whole story.

57%
GPQA Diamond
23%
Mock AIME
2,097,152
Context window
2024-09-24
Released
Google DeepMind
Developer
2
Price changes
2024-09-25
Tracked since
2.4× cheaper
Since first record

Recorded price changes

One row per change, newest first. US dollars per million tokens. Each row links to the exact commit or archived page the figure was read from.

DateInputOutput Cached inBlendedChangeSource
2024-11-07$1.25$5.00$2.19↓ 2.4×litellm@d0d29d7
2024-09-25$3.50$10.50$5.25litellm@5bc5eaf

How this was collected

Prices come from replaying the version history of LiteLLM's public pricing file and from vendor pricing pages archived by the Wayback Machine, then cleaned to remove metadata edits, same-day churn, unit slips and typo corrections. The full method and every script is in the repository. List pricing, not negotiated pricing.

See how the cost of a fixed level of capability has fallen →