Google: Gemini 3 Flash Preview✓ Catalog verified
Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool use performance with substantially lower latency than larger Gemini variants, making it well suited for interactive development, long running agent loops, and collaborative coding tasks. Compared to Gemini 2.5 Flash, it provides broad quality improvements across reasoning, multimodal understanding, and reliability. The model supports a 1M token context window and multimodal inputs including text, images, audio, video, and PDFs, with text output. It includes configurable reasoning via thinking levels (minimal, low, medium, high), structured output, tool use, and automatic context caching. Gemini 3 Flash Preview is optimized for users who want strong reasoning and agentic behavior without the cost or latency of full scale frontier models.
Overview
Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool use performance with substantially lower latency than larger Gemini variants, making it well suited for interactive development, long running agent loops, and collaborative coding tasks. Compared to Gemini 2.5 Flash, it provides broad quality improvements across reasoning, multimodal understanding, and reliability. The model supports a 1M token context window and multimodal inputs including text, images, audio, video, and PDFs, with text output. It includes configurable reasoning via thinking levels (minimal, low, medium, high), structured output, tool use, and automatic context caching. Gemini 3 Flash Preview is optimized for users who want strong reasoning and agentic behavior without the cost or latency of full scale frontier models.
Access: available through the official Google API. Context window: 1.0M.
Specifications
| API identifier | google/gemini-3-flash-preview |
| Provider | |
| Model type | Multimodal LLM |
| Context window | 1.0M catalog |
| Input modalities | text · image · file · audio · video |
| Output modalities | text |
| Released | Dec 17, 2025 |
| Tokenizer | Gemini |
| Moderated | No |
| Architecture modality | text+image+file+audio+video->text |
API defaults
{
"top_p": null,
"temperature": null,
"frequency_penalty": null
}Pricing
Live pricing components for google/gemini-3-flash-preview as published in the catalog.
| Pricing component | Raw unit price | Normalized | Unit |
|---|---|---|---|
| Prompt tokens | $5e-7 | $0.50 / 1M | Per input token |
| Completion tokens | $0.000003 | $3.00 / 1M | Per output token |
| Request fee | N/A | N/A | Per request |
| Image fee | $5e-7 | $0.000000 | Per image unit |
| Web search fee | N/A | N/A | Per search request |
Cost calculator
Estimate monthly spend from your own token volumes.
Estimated Total Cost
$0.85Calculated from current list pricing in the ModelsAtlas catalog.Capabilities
| Capability | Status | What it means |
|---|---|---|
| Visual Understanding | ✓ Supported | Image and document analysis support |
| Audio Processing | ✓ Supported | Speech and voice aligned flows |
| Tool Calling | ✓ Supported | Supports tools / function calling |
| Self-Hosting | Not advertised | Deploy outside managed APIs |
Capability signals
Directional signals derived from model metadata and capability tags — not official benchmark submissions.
Quick start
Call Google: Gemini 3 Flash Preview through an OpenAI-compatible client.
from openai import OpenAI
client = OpenAI(
base_url="https://openrouter.ai/api/v1",
api_key="YOUR_API_KEY"
)
response = client.chat.completions.create(
model="google/gemini-3-flash-preview",
messages=[{"role": "user", "content": "Explain quantum physics."}]
)
print(response.choices[0].message.content)Alternatives
Closest models by context window from a different provider.
| Model | Provider | Context | Input /1M | Actions |
|---|---|---|---|---|
| Meta: Llama 4 Maverick | Meta | 1.0M | $0.15 | View → |
| Xiaomi: MiMo-V2-Pro | Xiaomi | 1.0M | $1.00 | View → |
Head-to-head comparisons
Side-by-side breakdowns of Google: Gemini 3 Flash Preview against other frontier models.
Sources & attribution
Pricing and metadata are maintained in the ModelsAtlas catalog.
Frequently asked questions
How much does Google: Gemini 3 Flash Preview cost?
$0.50 per 1M input tokens and $3.00 per 1M output tokens on the official Google API.
What is the context window of Google: Gemini 3 Flash Preview?
Google: Gemini 3 Flash Preview supports up to 1.0M tokens of context.
Does Google: Gemini 3 Flash Preview support tool / function calling?
Yes, Google: Gemini 3 Flash Preview supports tool / function calling — you can register tools and the model will emit structured tool calls.
How do I access Google: Gemini 3 Flash Preview?
Use the official Google API with the model id `google/gemini-3-flash-preview`. See the Quick start section above for code examples.