Mistral Nemo 12B
Pricing & Specs
Mistral · $0.15 input / $0.15 output per 1M tokens · 128K context window
No longer offered on OpenMark
Mistral Nemo 12B is no longer offered on OpenMark (deprecated or superseded upstream). Its specs are kept here for reference — if you're still running it in production, it's worth benchmarking current alternatives: newer models are often cheaper AND better on the same task.
Mistral Nemo 12B API Pricing
| Tokens | Price |
|---|---|
| Input | $0.15 / 1M tokens |
| Output | $0.15 / 1M tokens |
50% discount available for batch API requests. Batch API pricing is available for this model at 50%. No cache pricing is published.
What that means in practice
Mistral Nemo 12B Specs
| Spec | Value |
|---|---|
| Context window | 128K tokens |
| Input modalities | text |
| Output modalities | text |
| Latency (measured by OpenMark) | ~218ms median response |
| Reasoning model | No |
| Tool / function calling | Yes |
| JSON mode | Yes |
| Streaming | Yes |
| Prompt caching | No |
| Batch API | Yes |
Notes
Retiring July 31, 2026. Replaced by ministral-8b-2512.
FAQ
How much does Mistral Nemo 12B cost?
Mistral Nemo 12B costs $0.15 per 1M input tokens and $0.15 per 1M output tokens.
What is Mistral Nemo 12B's context window?
Mistral Nemo 12B has a 128K-token context window.
Can I test Mistral Nemo 12B on my own task?
Mistral Nemo 12B is no longer offered on OpenMark, but you can benchmark its current alternatives from Mistral and other providers on your own task — free tier available.
Still using Mistral Nemo 12B?
It's been superseded. Benchmark its current alternatives on your actual
workload — you'll likely find something cheaper AND better. Free tier available.