llama-2-70b-chat-hf

deepinfra resold GPQA 26% 2.8× cheaper

Also listed as anyscale/meta-llama/Llama-2-70b-chat-hf, deepinfra/meta-llama/Llama-2-70b-chat-hf

Input$0.640per 1M tokens
Output$0.800per 1M tokens
Cached inputn/anot offered
A 1,000-in / 300-out request costs $0.0009 Reseller price · deepinfra · recorded 2025-08-21

Price over time

$0.500$1.00$2.002023-10-14today
26%
GPQA Diamond
0%
Mock AIME
4,096
Context window
2023-07-18
Released
Meta AI
Developer
4
Price changes
2023-10-14
Tracked since
2.8× cheaper
Since first record

Recorded price changes

One row per change, newest first. US dollars per million tokens. Each row links to the exact commit or archived page the figure was read from.

DateInputOutput Cached inBlendedChangeSource
2025-08-21$0.640$0.800$0.680↓ 1.1×litellm@9c4a86e
2024-01-08$0.700$0.900$0.750↓ 1.0×litellm@2486f92
2023-10-20$0.700$0.950$0.762↓ 2.5×litellm@a96de1f
2023-10-14$1.88$1.88$1.88litellm@8251e71

How this was collected

Prices come from replaying the version history of LiteLLM's public pricing file and from vendor pricing pages archived by the Wayback Machine, then cleaned to remove metadata edits, same-day churn, unit slips and typo corrections. The full method and every script is in the repository. List pricing, not negotiated pricing.

See how the cost of a fixed level of capability has fallen →