gpt-oss-20b

groq resold GPQA 61% 1.6× cheaper

Also listed as bedrock_mantle/openai.gpt-oss-20b, bedrock_mantle/us-gov-east-1/openai.gpt-oss-20b, bedrock_mantle/us-gov-west-1/openai.gpt-oss-20b, cerebras/openai/gpt-oss-20b, cloudflare/@cf/openai/gpt-oss-20b, darkbloom/gpt-oss-20b, databricks/databricks-gpt-oss-20b, deepinfra/openai/gpt-oss-20b, fireworks_ai/accounts/fireworks/models/gpt-oss-20b, fireworks_ai/gpt-oss-20b

Input$0.070per 1M tokens
Output$0.300per 1M tokens
Cached input$0.035per 1M tokens
A 1,000-in / 300-out request costs $0.0002 Reseller price · groq · recorded 2026-06-18

Price over time

$0.100$0.2002025-08-07today

The same capability, cheaper

gpt-oss-20b scores 61% on GPQA Diamond. The cheapest model clearing the GPQA 60%+ bar today is qwen3-7-flash at $0.055 blended — 2.3× less than this model's $0.128. Blended = 75% input, 25% output. Capability is one benchmark, not the whole story.

61%
GPQA Diamond
65%
Mock AIME
131,072
Context window
2025-08-05
Released
OpenAI
Developer
4
Price changes
2025-08-07
Tracked since
1.6× cheaper
Since first record

Recorded price changes

One row per change, newest first. US dollars per million tokens. Each row links to the exact commit or archived page the figure was read from.

DateInputOutput Cached inBlendedChangeSource
2026-06-18$0.070$0.300$0.035$0.128↓ 1.0×litellm@fb34c18
2026-01-19$0.075$0.300$0.037$0.131↑ 1.5×litellm@d48df6c
2025-09-06$0.050$0.200$0.088↓ 2.3×archived page
2025-08-07$0.100$0.500$0.200litellm@9c5e9d7

How this was collected

Prices come from replaying the version history of LiteLLM's public pricing file and from vendor pricing pages archived by the Wayback Machine, then cleaned to remove metadata edits, same-day churn, unit slips and typo corrections. The full method and every script is in the repository. List pricing, not negotiated pricing.

See how the cost of a fixed level of capability has fallen →