Gemini 3.6 Flash
Pricing & Specs
Google (Gemini) · $1.5 input / $7.5 output per 1M tokens · 1.04858M context window
The per-token price never tells the full story. A typical task (1K input + 500 output tokens) on Gemini 3.6 Flash costs about $0.0052, roughly $5.25 per 1,000 runs. But if it needs 3x the tokens of a cheaper model to match quality on YOUR task, the economics flip. The only way to know is to benchmark it on your actual workload.
Gemini 3.6 Flash API Pricing
| Tokens | Price |
|---|---|
| Input | $1.5 / 1M tokens |
| Output | $7.5 / 1M tokens |
| Cached input | $0.15 / 1M tokens |
Intro pricing $0.75/$3.75 (cache $0.075) through Dec 31, 2026; registry uses permanent Jan 1, 2027 price for benchmarking. Batch at 50%. Caching storage: $1.00 / 1M tokens per hour.
What that means in practice
Gemini 3.6 Flash Specs
| Spec | Value |
|---|---|
| Context window | 1.04858M tokens |
| Max output tokens | 65.536K tokens |
| Input modalities | text, image, video, audio, PDF |
| Output modalities | text |
| Reasoning model | Yes |
| Tool / function calling | Yes |
| JSON mode | Yes |
| Streaming | Yes |
| Prompt caching | Yes |
| Batch API | Yes |
Notes
Frontier-level intelligence at Flash speed and cost. Excels at code generation, agentic execution, and spatial reasoning.
FAQ
How much does Gemini 3.6 Flash cost?
Gemini 3.6 Flash costs $1.5 per 1M input tokens and $7.5 per 1M output tokens ($0.15 per 1M cached input tokens).
What is Gemini 3.6 Flash's context window?
Gemini 3.6 Flash has a 1.04858M-token context window. Maximum output is 65.536K tokens.
Can I test Gemini 3.6 Flash on my own task?
Yes. OpenMark lets you benchmark Gemini 3.6 Flash against 100+ models on your own task with real API calls, with no API keys needed, free tier available.
Is Gemini 3.6 Flash the right model for YOUR task?
Pricing tables can't answer that. Benchmark it against 100+ models
on your actual workload: real API calls, real costs. Free tier available.