Claude Haiku 3.5
Pricing & Specs
Anthropic · $0.8 input / $4 output per 1M tokens · 200K context window
No longer offered on OpenMark
Claude Haiku 3.5 is no longer offered on OpenMark (deprecated or superseded upstream). Its specs are kept here for reference — if you're still running it in production, it's worth benchmarking current alternatives: newer models are often cheaper AND better on the same task.
Claude Haiku 3.5 API Pricing
| Tokens | Price |
|---|---|
| Input | $0.8 / 1M tokens |
| Output | $4 / 1M tokens |
| Cached input | $0.08 / 1M tokens |
50% discount available for batch API requests. Batch API pricing is available for this model at 50%. Cache writes and cache hits or refreshes are billed separately.
What that means in practice
Claude Haiku 3.5 Specs
| Spec | Value |
|---|---|
| Context window | 200K tokens |
| Max output tokens | 8.192K tokens |
| Input modalities | text, image |
| Output modalities | text |
| Latency (measured by OpenMark) | ~584ms median response |
| Reasoning model | No |
| Tool / function calling | Yes |
| JSON mode | Yes |
| Streaming | Yes |
| Prompt caching | Yes |
| Batch API | Yes |
Capabilities
Notes
⚠️ Retiring Feb 19, 2026. Use claude-haiku-4 instead.
FAQ
How much does Claude Haiku 3.5 cost?
Claude Haiku 3.5 costs $0.8 per 1M input tokens and $4 per 1M output tokens ($0.08 per 1M cached input tokens).
What is Claude Haiku 3.5's context window?
Claude Haiku 3.5 has a 200K-token context window. Maximum output is 8.192K tokens.
Can I test Claude Haiku 3.5 on my own task?
Claude Haiku 3.5 is no longer offered on OpenMark, but you can benchmark its current alternatives from Anthropic and other providers on your own task — free tier available.
Still using Claude Haiku 3.5?
It's been superseded. Benchmark its current alternatives on your actual
workload — you'll likely find something cheaper AND better. Free tier available.