Cogito v2 preview 405B
Pricing & Specs
Deep Cogito · $3.5 input / $3.5 output per 1M tokens · 32.768K context window
No longer offered on OpenMark
Cogito v2 preview 405B is no longer offered on OpenMark (deprecated or superseded upstream). Its specs are kept here for reference — if you're still running it in production, it's worth benchmarking current alternatives: newer models are often cheaper AND better on the same task.
Cogito v2 preview 405B API Pricing
| Tokens | Price |
|---|---|
| Input | $3.5 / 1M tokens |
| Output | $3.5 / 1M tokens |
Batch API pricing is available via Together at 50%; cache pricing is not published.
What that means in practice
Cogito v2 preview 405B Specs
| Spec | Value |
|---|---|
| Context window | 32.768K tokens |
| Input modalities | text |
| Output modalities | text |
| Latency (measured by OpenMark) | ~1.9s median response |
| Reasoning model | No |
| Tool / function calling | No |
| JSON mode | Yes |
| Streaming | Yes |
| Prompt caching | No |
| Batch API | Yes |
Capabilities
Notes
⚠️ Retiring Feb 4, 2026. Use cogito-v2-1-671b instead.
FAQ
How much does Cogito v2 preview 405B cost?
Cogito v2 preview 405B costs $3.5 per 1M input tokens and $3.5 per 1M output tokens.
What is Cogito v2 preview 405B's context window?
Cogito v2 preview 405B has a 32.768K-token context window.
Can I test Cogito v2 preview 405B on my own task?
Cogito v2 preview 405B is no longer offered on OpenMark, but you can benchmark its current alternatives from Deep Cogito and other providers on your own task — free tier available.
Still using Cogito v2 preview 405B?
It's been superseded. Benchmark its current alternatives on your actual
workload — you'll likely find something cheaper AND better. Free tier available.