Grok 3 Mini Fast
Pricing & Specs
xAI · $0.6 input / $4 output per 1M tokens · 131.072K context window
No longer offered on OpenMark
Grok 3 Mini Fast is no longer offered on OpenMark (deprecated or superseded upstream). Its specs are kept here for reference. If you're still running it in production, it's worth benchmarking current alternatives: newer models are often cheaper AND better on the same task.
Grok 3 Mini Fast API Pricing
| Tokens | Price |
|---|---|
| Input | $0.6 / 1M tokens |
| Output | $4 / 1M tokens |
| Cached input | $0.15 / 1M tokens |
Cached tokens pricing is available for this model.
What that means in practice
Grok 3 Mini Fast Specs
| Spec | Value |
|---|---|
| Context window | 131.072K tokens |
| Input modalities | text |
| Output modalities | text |
| Latency (measured by OpenMark) | ~2.0s median response |
| Reasoning model | Yes |
| Tool / function calling | Yes |
| JSON mode | Yes |
| Streaming | Yes |
| Prompt caching | Yes |
Capabilities
Notes
Removed from xAI docs Jan 2026. Use grok-4-fast-non-reasoning instead.
FAQ
How much does Grok 3 Mini Fast cost?
Grok 3 Mini Fast costs $0.6 per 1M input tokens and $4 per 1M output tokens ($0.15 per 1M cached input tokens).
What is Grok 3 Mini Fast's context window?
Grok 3 Mini Fast has a 131.072K-token context window.
Can I test Grok 3 Mini Fast on my own task?
Grok 3 Mini Fast is no longer offered on OpenMark, but you can benchmark its current alternatives from xAI and other providers on your own task, free tier available.
Still using Grok 3 Mini Fast?
It's been superseded. Benchmark its current alternatives on your actual
workload. You'll likely find something cheaper AND better. Free tier available.