Token Price History / together_ai

qwen3-5-397b-a17b

together_ai resold GPQA 86%

Also listed as deepinfra/Qwen/Qwen3.5-397B-A17B, nebius/Qwen/Qwen3.5-397B-A17B, novita/qwen/qwen3.5-397b-a17b, openrouter/qwen/qwen3.5-397b-a17b, scaleway/qwen/qwen3.5-397b-a17b, together_ai/Qwen/Qwen/Qwen3.5-397B-A17B, together_ai/Qwen/Qwen3.5-397B-A17B

Input$0.550per 1M tokens
Output$3.50per 1M tokens
Cached inputn/anot offered
A 1,000-in / 300-out request costs $0.0016 Reseller price · together_ai · recorded 2026-09-03

Price over time

$1.002025-12-15today

The same capability, cheaper

qwen3-5-397b-a17b scores 86% on GPQA Diamond. The cheapest model clearing the GPQA 80%+ bar today is qwen3-7-flash at $0.055 blended — 23.4× less than this model's $1.29. Blended = 75% input, 25% output. Capability is one benchmark, not the whole story.

86%
GPQA Diamond
89%
Mock AIME
262,144
Context window
2026-02-13
Released
Alibaba
Developer
3
Price changes
2025-12-15
Tracked since

Recorded price changes

One row per change, newest first. US dollars per million tokens. Each row links to the exact commit or archived page the figure was read from.

DateInputOutput Cached inBlendedChangeSource
2026-09-03$0.550$3.50$1.29↓ 1.0×litellm@2c4eb69
2026-08-26$0.600$3.60$0.350$1.35litellm@4456a44
2025-12-15$0.600$3.60$1.35litellm@20a41a6

How this was collected

Prices come from replaying the version history of LiteLLM's public pricing file and from vendor pricing pages archived by the Wayback Machine, then cleaned to remove metadata edits, same-day churn, unit slips and typo corrections. The full method and every script is in the repository. List pricing, not negotiated pricing.

See how the cost of a fixed level of capability has fallen →