Claude Opus 4.8
Pricing & Specs
Anthropic · $5 input / $25 output per 1M tokens · 200K context window
The per-token price never tells the full story. A typical task (1K input + 500 output tokens) on Claude Opus 4.8 costs about $0.0175, roughly $17.50 per 1,000 runs. But if it needs 3x the tokens of a cheaper model to match quality on YOUR task, the economics flip. The only way to know is to benchmark it on your actual workload.
Claude Opus 4.8 API Pricing
| Tokens | Price |
|---|---|
| Input | $5 / 1M tokens |
| Output | $25 / 1M tokens |
| Cached input | $0.5 / 1M tokens |
50% discount available for batch API requests. Batch API pricing is available for this model at 50%. Cache writes and cache hits or refreshes are billed separately.
What that means in practice
Claude Opus 4.8 Specs
| Spec | Value |
|---|---|
| Context window | 200K tokens |
| Max output tokens | 128K tokens |
| Input modalities | text, image |
| Output modalities | text |
| Knowledge cutoff | Aug 2025 |
| Reasoning model | Yes |
| Tool / function calling | Yes |
| JSON mode | Yes |
| Streaming | Yes |
| Prompt caching | Yes |
| Batch API | Yes |
Notes
No hidden reasoning by default, but reasoning can still occur in output tokens. OpenMark uses out-of-the-box default parameters.
FAQ
How much does Claude Opus 4.8 cost?
Claude Opus 4.8 costs $5 per 1M input tokens and $25 per 1M output tokens ($0.5 per 1M cached input tokens).
What is Claude Opus 4.8's context window?
Claude Opus 4.8 has a 200K-token context window. Maximum output is 128K tokens.
Can I test Claude Opus 4.8 on my own task?
Yes. OpenMark lets you benchmark Claude Opus 4.8 against 100+ models on your own task with real API calls, with no API keys needed, free tier available.
Is Claude Opus 4.8 the right model for YOUR task?
Pricing tables can't answer that. Benchmark it against 100+ models
on your actual workload: real API calls, real costs. Free tier available.