Kimi K2
Pricing & Specs
Moonshot AI · $0.5 input / $2.8 output per 1M tokens · 262.144K context window
No longer offered on OpenMark
Kimi K2 is no longer offered on OpenMark (deprecated or superseded upstream). Its specs are kept here for reference — if you're still running it in production, it's worth benchmarking current alternatives: newer models are often cheaper AND better on the same task.
Kimi K2 API Pricing
| Tokens | Price |
|---|---|
| Input | $0.5 / 1M tokens |
| Output | $2.8 / 1M tokens |
25% discount available for batch API requests.
What that means in practice
Kimi K2 Specs
| Spec | Value |
|---|---|
| Context window | 262.144K tokens |
| Input modalities | text, image |
| Output modalities | text |
| Latency (measured by OpenMark) | ~700ms median response |
| Reasoning model | No |
| Tool / function calling | Yes |
| JSON mode | Yes |
| Streaming | Yes |
| Prompt caching | No |
| Batch API | Yes |
Capabilities
Notes
Scheduled for deprecation on Together AI: March 06, 2026.
FAQ
How much does Kimi K2 cost?
Kimi K2 costs $0.5 per 1M input tokens and $2.8 per 1M output tokens.
What is Kimi K2's context window?
Kimi K2 has a 262.144K-token context window.
Can I test Kimi K2 on my own task?
Kimi K2 is no longer offered on OpenMark, but you can benchmark its current alternatives from Moonshot AI and other providers on your own task — free tier available.
Still using Kimi K2?
It's been superseded. Benchmark its current alternatives on your actual
workload — you'll likely find something cheaper AND better. Free tier available.