Google: Gemini 2.5 Flash✓ Catalog verified
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater accuracy and nuanced context handling. Additionally, Gemini 2.5 Flash is configurable through the "max tokens for reasoning" parameter, as described in the documentation (https://openrouter.ai/docs/use-cases/reasoning-tokens#max-tokens-for-reasoning).
Overview
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater accuracy and nuanced context handling. Additionally, Gemini 2.5 Flash is configurable through the "max tokens for reasoning" parameter, as described in the documentation (https://openrouter.ai/docs/use-cases/reasoning-tokens#max-tokens-for-reasoning).
Access: available through the official Google API. Context window: 1.0M.
Specifications
| API identifier | google/gemini-2.5-flash |
| Provider | |
| Model type | Multimodal LLM |
| Context window | 1.0M catalog |
| Input modalities | file · image · text · audio · video |
| Output modalities | text |
| Released | Jun 17, 2025 |
| Tokenizer | Gemini |
| Moderated | No |
| Architecture modality | text+image+file+audio+video->text |
API defaults
{
"top_p": null,
"temperature": null,
"frequency_penalty": null
}Pricing
Live pricing components for google/gemini-2.5-flash as published in the catalog.
| Pricing component | Raw unit price | Normalized | Unit |
|---|---|---|---|
| Prompt tokens | $3e-7 | $0.30 / 1M | Per input token |
| Completion tokens | $0.0000025 | $2.50 / 1M | Per output token |
| Request fee | N/A | N/A | Per request |
| Image fee | $3e-7 | $0.000000 | Per image unit |
| Web search fee | N/A | N/A | Per search request |
Cost calculator
Estimate monthly spend from your own token volumes.
Estimated Total Cost
$0.65Calculated from current list pricing in the ModelsAtlas catalog.Capabilities
| Capability | Status | What it means |
|---|---|---|
| Visual Understanding | ✓ Supported | Image and document analysis support |
| Audio Processing | ✓ Supported | Speech and voice aligned flows |
| Tool Calling | ✓ Supported | Supports tools / function calling |
| Self-Hosting | Not advertised | Deploy outside managed APIs |
Capability signals
Directional signals derived from model metadata and capability tags — not official benchmark submissions.
Quick start
Call Google: Gemini 2.5 Flash through an OpenAI-compatible client.
from openai import OpenAI
client = OpenAI(
base_url="https://openrouter.ai/api/v1",
api_key="YOUR_API_KEY"
)
response = client.chat.completions.create(
model="google/gemini-2.5-flash",
messages=[{"role": "user", "content": "Explain quantum physics."}]
)
print(response.choices[0].message.content)Alternatives
Closest models by context window from a different provider.
| Model | Provider | Context | Input /1M | Actions |
|---|---|---|---|---|
| Meta: Llama 4 Maverick | Meta | 1.0M | $0.15 | View → |
| Xiaomi: MiMo-V2-Pro | Xiaomi | 1.0M | $1.00 | View → |
Head-to-head comparisons
Side-by-side breakdowns of Google: Gemini 2.5 Flash against other frontier models.
Sources & attribution
Pricing and metadata are maintained in the ModelsAtlas catalog.
Frequently asked questions
How much does Google: Gemini 2.5 Flash cost?
$0.30 per 1M input tokens and $2.50 per 1M output tokens on the official Google API.
What is the context window of Google: Gemini 2.5 Flash?
Google: Gemini 2.5 Flash supports up to 1.0M tokens of context.
Does Google: Gemini 2.5 Flash support tool / function calling?
Yes, Google: Gemini 2.5 Flash supports tool / function calling — you can register tools and the model will emit structured tool calls.
How do I access Google: Gemini 2.5 Flash?
Use the official Google API with the model id `google/gemini-2.5-flash`. See the Quick start section above for code examples.