Atlas / Models / Google / Google: Gemini 3 Flash Preview

Google: Gemini 3 Flash Preview✓ Catalog verified

google/gemini-3-flash-preview

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool use performance with substantially lower latency than larger Gemini variants, making it well suited for interactive development, long running agent loops, and collaborative coding tasks. Compared to Gemini 2.5 Flash, it provides broad quality improvements across reasoning, multimodal understanding, and reliability. The model supports a 1M token context window and multimodal inputs including text, images, audio, video, and PDFs, with text output. It includes configurable reasoning via thinking levels (minimal, low, medium, high), structured output, tool use, and automatic context caching. Gemini 3 Flash Preview is optimized for users who want strong reasoning and agentic behavior without the cost or latency of full scale frontier models.

Input price
$0.50 /1M
Output price
$3.00 /1M
Context
1.0M
Modalities
textimagefileaudio
Released
Dec 17, 2025
Tool calling
✓ Yes
Atlas signal
91/100
01

Overview

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool use performance with substantially lower latency than larger Gemini variants, making it well suited for interactive development, long running agent loops, and collaborative coding tasks. Compared to Gemini 2.5 Flash, it provides broad quality improvements across reasoning, multimodal understanding, and reliability. The model supports a 1M token context window and multimodal inputs including text, images, audio, video, and PDFs, with text output. It includes configurable reasoning via thinking levels (minimal, low, medium, high), structured output, tool use, and automatic context caching. Gemini 3 Flash Preview is optimized for users who want strong reasoning and agentic behavior without the cost or latency of full scale frontier models.

Access: available through the official Google API. Context window: 1.0M.

Pricing and metadata from the ModelsAtlas catalog.Last refreshed Aug 5, 2026
02

Specifications

API identifiergoogle/gemini-3-flash-preview
ProviderGoogle
Model typeMultimodal LLM
Context window1.0M catalog
Input modalitiestext · image · file · audio · video
Output modalitiestext
ReleasedDec 17, 2025
TokenizerGemini
ModeratedNo
Architecture modalitytext+image+file+audio+video->text
Supported parameters
include_reasoningmax_tokensreasoningresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_p
Source: provider documentation + catalog feedMethodology →

API defaults

Default parameters
{
  "top_p": null,
  "temperature": null,
  "frequency_penalty": null
}
03

Pricing

Live pricing components for google/gemini-3-flash-preview as published in the catalog.

$0.50
Input /1M
$3.00
Output /1M
Per request
$0.000000
Per image
Pricing componentRaw unit priceNormalizedUnit
Prompt tokens$5e-7$0.50 / 1MPer input token
Completion tokens$0.000003$3.00 / 1MPer output token
Request feeN/AN/APer request
Image fee$5e-7$0.000000Per image unit
Web search feeN/AN/APer search request
Source: catalog pricing feedCompare all pricing →Cheapest models →
04

Cost calculator

Estimate monthly spend from your own token volumes.

Scale / Volume

Estimated Total Cost

$0.85Calculated from current list pricing in the ModelsAtlas catalog.
05

Capabilities

CapabilityStatusWhat it means
Visual Understanding✓ SupportedImage and document analysis support
Audio Processing✓ SupportedSpeech and voice aligned flows
Tool Calling✓ SupportedSupports tools / function calling
Self-HostingNot advertisedDeploy outside managed APIs
Derived from catalog capability tags and supported parameters
06

Capability signals

Directional signals derived from model metadata and capability tags — not official benchmark submissions.

Google: Gemini 3 Flash Preview
MMLU Signal (Reasoning)
97signal
Coding Signal (HumanEval proxy)
92signal
Math Signal (GSM8K proxy)
87signal
Science Signal (GPQA proxy)
87signal
Estimated from metadata — not an official benchmark runAll benchmarks →Methodology →
07

Quick start

Call Google: Gemini 3 Flash Preview through an OpenAI-compatible client.

from openai import OpenAI

client = OpenAI(
    base_url="https://openrouter.ai/api/v1",
    api_key="YOUR_API_KEY"
)

response = client.chat.completions.create(
    model="google/gemini-3-flash-preview",
    messages=[{"role": "user", "content": "Explain quantum physics."}]
)

print(response.choices[0].message.content)
08

Alternatives

Closest models by context window from a different provider.

09

Sources & attribution

Pricing and metadata are maintained in the ModelsAtlas catalog.

Capability bars use metadata tags — directional estimates onlyHow we source and verify data →
10

Frequently asked questions

How much does Google: Gemini 3 Flash Preview cost?

$0.50 per 1M input tokens and $3.00 per 1M output tokens on the official Google API.

What is the context window of Google: Gemini 3 Flash Preview?

Google: Gemini 3 Flash Preview supports up to 1.0M tokens of context.

Does Google: Gemini 3 Flash Preview support tool / function calling?

Yes, Google: Gemini 3 Flash Preview supports tool / function calling — you can register tools and the model will emit structured tool calls.

How do I access Google: Gemini 3 Flash Preview?

Use the official Google API with the model id `google/gemini-3-flash-preview`. See the Quick start section above for code examples.