Codestral 2508
Pricing & Specs
Mistral · $0.3 input / $0.9 output per 1M tokens · 128K context window
The per-token price never tells the full story. A typical task (1K input + 500 output tokens) on Codestral 2508 costs about $0.0008, roughly $0.75 per 1,000 runs. But if it needs 3x the tokens of a cheaper model to match quality on YOUR task, the economics flip. The only way to know is to benchmark it on your actual workload.
Codestral 2508 API Pricing
| Tokens | Price |
|---|---|
| Input | $0.3 / 1M tokens |
| Output | $0.9 / 1M tokens |
50% discount available for batch API requests. Batch API pricing is available for this model at 50%. No cache pricing is published.
What that means in practice
Codestral 2508 Specs
| Spec | Value |
|---|---|
| Context window | 128K tokens |
| Input modalities | text |
| Output modalities | text |
| Latency (measured by OpenMark) | ~132ms median response |
| Reasoning model | No |
| Tool / function calling | Yes |
| JSON mode | Yes |
| Streaming | Yes |
| Prompt caching | No |
| Batch API | Yes |
Capabilities
FAQ
How much does Codestral 2508 cost?
Codestral 2508 costs $0.3 per 1M input tokens and $0.9 per 1M output tokens.
What is Codestral 2508's context window?
Codestral 2508 has a 128K-token context window.
Can I test Codestral 2508 on my own task?
Yes. OpenMark lets you benchmark Codestral 2508 against 100+ models on your own task with real API calls, with no API keys needed, free tier available.
Is Codestral 2508 the right model for YOUR task?
Pricing tables can't answer that. Benchmark it against 100+ models
on your actual workload: real API calls, real costs. Free tier available.