Gemini 2.5 Flash
Pricing & Specs

Google (Gemini) · $0.3 input / $2.5 output per 1M tokens · 1.04858M context window

$0.3
per 1M input tokens
$2.5
per 1M output tokens
1.04858M
context window (tokens)

The per-token price never tells the full story. A typical task (1K input + 500 output tokens) on Gemini 2.5 Flash costs about $0.0015 — roughly $1.55 per 1,000 runs. But if it needs 3x the tokens of a cheaper model to match quality on YOUR task, the economics flip. The only way to know is to benchmark it on your actual workload.

Gemini 2.5 Flash API Pricing

TokensPrice
Input$0.3 / 1M tokens
Output$2.5 / 1M tokens
Cached input$0.03 / 1M tokens

Context caching tiers: $0.075, $0.25, $1.0; Context caching storage: $1.00 / 1,000,000 tokens per hour

What that means in practice

1x One typical task (1K input + 500 output tokens): $0.0015
1K 1,000 runs of that task: $1.55
💡 Real cost depends on how verbose the model is on YOUR prompts — benchmark to measure actual tokens, not estimates.

Gemini 2.5 Flash Specs

SpecValue
Context window1.04858M tokens
Max output tokens65.536K tokens
Input modalitiestext, image, audio, video
Output modalitiestext
Latency (measured by OpenMark)~605ms median response
Reasoning modelYes
Tool / function callingYes
JSON modeYes
StreamingYes
Prompt cachingYes
Batch APIYes

Capabilities

text generation vision function calling structured outputs streaming code execution search grounding thinking

FAQ

How much does Gemini 2.5 Flash cost?

Gemini 2.5 Flash costs $0.3 per 1M input tokens and $2.5 per 1M output tokens ($0.03 per 1M cached input tokens).

What is Gemini 2.5 Flash's context window?

Gemini 2.5 Flash has a 1.04858M-token context window. Maximum output is 65.536K tokens.

Can I test Gemini 2.5 Flash on my own task?

Yes. OpenMark lets you benchmark Gemini 2.5 Flash against 100+ models on your own task with real API calls — no API keys needed, free tier available.

Is Gemini 2.5 Flash the right model for YOUR task?

Pricing tables can't answer that. Benchmark it against 100+ models
on your actual workload — real API calls, real costs. Free tier available.

Benchmark Gemini 2.5 Flash — Free →