Qwen 2 72B Instruct API
Pricing & Specs
Qwen (Alibaba) · $0.9 input / $0.9 output per 1M tokens · 32.768K context window
No longer offered on OpenMark
Qwen 2 72B Instruct API is no longer offered on OpenMark (deprecated or superseded upstream). Its specs are kept here for reference. If you're still running it in production, it's worth benchmarking current alternatives: newer models are often cheaper AND better on the same task.
Qwen 2 72B Instruct API API Pricing
| Tokens | Price |
|---|---|
| Input | $0.9 / 1M tokens |
| Output | $0.9 / 1M tokens |
Batch API pricing is available via Together at 50%; cache pricing is not published.
What that means in practice
Qwen 2 72B Instruct API Specs
| Spec | Value |
|---|---|
| Context window | 32.768K tokens |
| Input modalities | text |
| Output modalities | text |
| Reasoning model | No |
| Tool / function calling | No |
| JSON mode | No |
| Streaming | Yes |
| Prompt caching | No |
| Batch API | Yes |
Capabilities
Notes
Removed by host on 2025-08
FAQ
How much does Qwen 2 72B Instruct API cost?
Qwen 2 72B Instruct API costs $0.9 per 1M input tokens and $0.9 per 1M output tokens.
What is Qwen 2 72B Instruct API's context window?
Qwen 2 72B Instruct API has a 32.768K-token context window.
Can I test Qwen 2 72B Instruct API on my own task?
Qwen 2 72B Instruct API is no longer offered on OpenMark, but you can benchmark its current alternatives from Qwen (Alibaba) and other providers on your own task, free tier available.
Still using Qwen 2 72B Instruct API?
It's been superseded. Benchmark its current alternatives on your actual
workload. You'll likely find something cheaper AND better. Free tier available.