Google: Gemini 2.0 Flash Lite✓ Catalog verified
Gemini 2.0 Flash Lite offers a significantly faster time to first token (TTFT) compared to Gemini Flash 1.5, while maintaining quality on par with larger models like Gemini Pro 1.5, all at extremely economical token prices.
Overview
Gemini 2.0 Flash Lite offers a significantly faster time to first token (TTFT) compared to Gemini Flash 1.5, while maintaining quality on par with larger models like Gemini Pro 1.5, all at extremely economical token prices.
Access: available through the official Google API. Context window: 1.0M.
Specifications
| API identifier | google/gemini-2.0-flash-lite-001 |
| Provider | |
| Model type | Multimodal LLM |
| Context window | 1.0M catalog |
| Input modalities | text · image · file · audio · video |
| Output modalities | text |
| Released | Feb 25, 2025 |
| Tokenizer | Gemini |
| Moderated | No |
| Architecture modality | text+image+file+audio+video->text |
| Scheduled expiry | 2026-06-01 |
API defaults
{
"top_p": null,
"temperature": null,
"frequency_penalty": null
}Pricing
Live pricing components for google/gemini-2.0-flash-lite-001 as published in the catalog.
| Pricing component | Raw unit price | Normalized | Unit |
|---|---|---|---|
| Prompt tokens | $7.5e-8 | $0.07 / 1M | Per input token |
| Completion tokens | $3e-7 | $0.30 / 1M | Per output token |
| Request fee | N/A | N/A | Per request |
| Image fee | $7.5e-8 | $0.000000 | Per image unit |
| Web search fee | N/A | N/A | Per search request |
Cost calculator
Estimate monthly spend from your own token volumes.
Estimated Total Cost
$0.1Calculated from current list pricing in the ModelsAtlas catalog.Capabilities
| Capability | Status | What it means |
|---|---|---|
| Visual Understanding | ✓ Supported | Image and document analysis support |
| Audio Processing | ✓ Supported | Speech and voice aligned flows |
| Tool Calling | ✓ Supported | Supports tools / function calling |
| Self-Hosting | Not advertised | Deploy outside managed APIs |
Capability signals
Directional signals derived from model metadata and capability tags — not official benchmark submissions.
Quick start
Call Google: Gemini 2.0 Flash Lite through an OpenAI-compatible client.
from openai import OpenAI
client = OpenAI(
base_url="https://openrouter.ai/api/v1",
api_key="YOUR_API_KEY"
)
response = client.chat.completions.create(
model="google/gemini-2.0-flash-lite-001",
messages=[{"role": "user", "content": "Explain quantum physics."}]
)
print(response.choices[0].message.content)Alternatives
Closest models by context window from a different provider.
| Model | Provider | Context | Input /1M | Actions |
|---|---|---|---|---|
| Meta: Llama 4 Maverick | Meta | 1.0M | $0.15 | View → |
| Xiaomi: MiMo-V2-Pro | Xiaomi | 1.0M | $1.00 | View → |
Sources & attribution
Pricing and metadata are maintained in the ModelsAtlas catalog.
Frequently asked questions
How much does Google: Gemini 2.0 Flash Lite cost?
$0.07 per 1M input tokens and $0.30 per 1M output tokens on the official Google API.
What is the context window of Google: Gemini 2.0 Flash Lite?
Google: Gemini 2.0 Flash Lite supports up to 1.0M tokens of context.
Does Google: Gemini 2.0 Flash Lite support tool / function calling?
Yes, Google: Gemini 2.0 Flash Lite supports tool / function calling — you can register tools and the model will emit structured tool calls.
How do I access Google: Gemini 2.0 Flash Lite?
Use the official Google API with the model id `google/gemini-2.0-flash-lite-001`. See the Quick start section above for code examples.