Private beta DPDP-ready sovereign routing is live, request access
Together

qwen-2.5-72b-turbo pricing

List price for the qwen-2.5-72b-turbo API from Together — input cost, output cost, and context window, plus what it actually costs to run across a few common workloads. All figures in USD per 1M tokens.

Last updated: July 2026
Input
$1.20 / 1M tokens
Output
$1.20 / 1M tokens
1M in + 1M out
$2.40 blended
Context window
128K tokens

What qwen-2.5-72b-turbo costs for typical workloads

Estimated cost per request and per 1,000 requests at qwen-2.5-72b-turbo's list price, for a few representative input/output token mixes. Your real usage will vary with prompt and response length.

Workload Input tokens Output tokens Cost / request Cost / 1,000
Short chat turn 1,000 300 $0.00156 $1.56
RAG answer 8,000 500 $0.010 $10.20
Agent step (tool-heavy) 20,000 2,000 $0.026 $26.40
Long-document summary 100,000 1,000 $0.121 $121

Provider-published list price, USD per 1M tokens, current as of July 2026. Batch, cached-input, and volume discounts are not applied. Verify with Together before relying on these figures for billing.

Route to the cheapest model automatically.

Stop hand-tuning which model gets which request. Routeplane's difficulty router scores each prompt and picks the optimal model per request — downgrading easy calls to a cheaper model and reserving the expensive ones for the hard prompts, with automatic fallback if a provider is down. One OpenAI-compatible endpoint, every provider behind it.