Qwen 2.5 72B Instruct Turbo API
Pricing & Specs
Qwen (Alibaba) · $1.2 input / $1.2 output per 1M tokens · 131.072K context window
The per-token price never tells the full story. A typical task (1K input + 500 output tokens) on Qwen 2.5 72B Instruct Turbo API costs about $0.0018 — roughly $1.80 per 1,000 runs. But if it needs 3x the tokens of a cheaper model to match quality on YOUR task, the economics flip. The only way to know is to benchmark it on your actual workload.
Qwen 2.5 72B Instruct Turbo API API Pricing
| Tokens | Price |
|---|---|
| Input | $1.2 / 1M tokens |
| Output | $1.2 / 1M tokens |
50% discount available for batch API requests. Batch API pricing is available via Together at 50%; cache pricing is not published.
What that means in practice
Qwen 2.5 72B Instruct Turbo API Specs
| Spec | Value |
|---|---|
| Context window | 131.072K tokens |
| Input modalities | text |
| Output modalities | text |
| Latency (measured by OpenMark) | ~237ms median response |
| Reasoning model | No |
| Tool / function calling | Yes |
| JSON mode | Yes |
| Streaming | Yes |
| Prompt caching | No |
| Batch API | Yes |
Capabilities
FAQ
How much does Qwen 2.5 72B Instruct Turbo API cost?
Qwen 2.5 72B Instruct Turbo API costs $1.2 per 1M input tokens and $1.2 per 1M output tokens.
What is Qwen 2.5 72B Instruct Turbo API's context window?
Qwen 2.5 72B Instruct Turbo API has a 131.072K-token context window.
Can I test Qwen 2.5 72B Instruct Turbo API on my own task?
Yes. OpenMark lets you benchmark Qwen 2.5 72B Instruct Turbo API against 100+ models on your own task with real API calls — no API keys needed, free tier available.
Is Qwen 2.5 72B Instruct Turbo API the right model for YOUR task?
Pricing tables can't answer that. Benchmark it against 100+ models
on your actual workload — real API calls, real costs. Free tier available.