Gemini 2.0 Flash-Lite
Pricing & Specs
Google (Gemini) · $0.075 input / $0.3 output per 1M tokens · 1.04858M context window
No longer offered on OpenMark
Gemini 2.0 Flash-Lite is no longer offered on OpenMark (deprecated or superseded upstream). Its specs are kept here for reference. If you're still running it in production, it's worth benchmarking current alternatives: newer models are often cheaper AND better on the same task.
Gemini 2.0 Flash-Lite API Pricing
| Tokens | Price |
|---|---|
| Input | $0.075 / 1M tokens |
| Output | $0.3 / 1M tokens |
Context caching is available and storage fees may apply. Batch API pricing is available for this model at 50%. Some models use prompt-size thresholds that affect rates.
What that means in practice
Gemini 2.0 Flash-Lite Specs
| Spec | Value |
|---|---|
| Context window | 1.04858M tokens |
| Max output tokens | 8.192K tokens |
| Input modalities | text, image, audio, video |
| Output modalities | text |
| Latency (measured by OpenMark) | ~417ms median response |
| Reasoning model | No |
| Tool / function calling | Yes |
| JSON mode | Yes |
| Streaming | Yes |
| Prompt caching | No |
| Batch API | Yes |
Capabilities
FAQ
How much does Gemini 2.0 Flash-Lite cost?
Gemini 2.0 Flash-Lite costs $0.075 per 1M input tokens and $0.3 per 1M output tokens.
What is Gemini 2.0 Flash-Lite's context window?
Gemini 2.0 Flash-Lite has a 1.04858M-token context window. Maximum output is 8.192K tokens.
Can I test Gemini 2.0 Flash-Lite on my own task?
Gemini 2.0 Flash-Lite is no longer offered on OpenMark, but you can benchmark its current alternatives from Google (Gemini) and other providers on your own task, free tier available.
Still using Gemini 2.0 Flash-Lite?
It's been superseded. Benchmark its current alternatives on your actual
workload. You'll likely find something cheaper AND better. Free tier available.