Google: Gemini 2.5 Flash Lite Preview 09-2025 vs MiniMax: MiniMax M3 (free)

Neither wins outright: Google: Gemini 2.5 Flash Lite Preview 09-2025 is the pick when you need audio input, MiniMax: MiniMax M3 (free) when cost matters at scale.

Key numbers

SpecGoogle: Gemini 2.5 Flash Lite Preview 09-2025MiniMax: MiniMax M3 (free)
ProviderGoogleMinimax
Input price (per 1M tokens)$0.10$0
Output price (per 1M tokens)$0.40$0
Blended price (3:1 input/output)$0.18$0
Cheaper than … of all models we track72%90%
Context window1.0M tokens (~1,573 pages)1.0M tokens (~1,573 pages)
Larger context than … of all models81%81%
Max output per response66K tokens (~49,200 words)944K tokens (~707,800 words)
Input typestext, image, file, audio, videotext, image, video
Output typestexttext
Free variantNoYes
Listed since2025-09-252026-05-31

What it costs for real workloads

Monthly cost at list price for five common workloads. Token counts per unit are stated so you can scale them to your own traffic. A workload that does not fit a model's context window is marked rather than priced.

Workload (per month)Google: Gemini 2.5 Flash Lite Preview 09-2025MiniMax: MiniMax M3 (free)Difference
Customer-support chatbot
50,000 conversations — each a 6-turn conversation (1,500 in / 400 out tokens)
$15.50$0MiniMax: MiniMax M3 (free) saves $15.50
RAG search over documents
20,000 questions — each a question answered from ~8 retrieved passages (6,000 in / 500 out tokens)
$16.00$0MiniMax: MiniMax M3 (free) saves $16.00
Coding / tool-using agent
2,000 tasks — each a multi-step task with tool calls and file context (60,000 in / 6,000 out tokens)
$16.80$0MiniMax: MiniMax M3 (free) saves $16.80
Batch summarisation
10,000 documents — each a 6,000-word report summarised to a page (8,000 in / 600 out tokens)
$10.40$0MiniMax: MiniMax M3 (free) saves $10.40
Long-document analysis
500 documents — each a 150-page contract or codebase read in one pass (100,000 in / 2,000 out tokens)
$5.40$0MiniMax: MiniMax M3 (free) saves $5.40

Capabilities side by side

CapabilityGoogle: Gemini 2.5 Flash Lite Preview 09-2025MiniMax: MiniMax M3 (free)What it lets you do
Tool / function calling✓ Yes✓ Yescall your APIs and run agent loops
Structured output (JSON schema)✓ Yes✓ Yesreturn JSON that validates against your schema
Reasoning / thinking mode✓ Yes✓ Yesspend extra tokens thinking before answering hard problems
Image input✓ Yes✓ Yesread screenshots, charts and scanned pages
Audio input✓ Yes—take speech or audio directly
File / PDF input✓ Yes—accept documents without your own parsing step
Image output——generate images, not just text
Built-in web search——answer from live web results
Seed (reproducible sampling)✓ Yes✓ Yesrepeat a generation for tests and evals
Log-probabilities——read token confidence for classification and scoring

From each model's published parameters and input/output types. “—” means not advertised, not proven absent.

Price history

Which should you choose?

Choose Google: Gemini 2.5 Flash Lite Preview 09-2025 if…

  • you need audio input — to take speech or audio directly; MiniMax: MiniMax M3 (free) does not advertise it
  • you need file / PDF input — to accept documents without your own parsing step; MiniMax: MiniMax M3 (free) does not advertise it

Choose MiniMax: MiniMax M3 (free) if…

  • cost matters at scale — it is 100% cheaper on a typical 3:1 input/output mix ($0 vs $15.50 a month for 50,000 support chats)
  • you generate long outputs in one go — up to 944K tokens (about 707,800 words) per response
  • you want to prototype at zero cost — a free variant is listed
  • you want the newer model — released 2026-05-31, 8 months after the other

Frequently asked questions

Is Google: Gemini 2.5 Flash Lite Preview 09-2025 better than MiniMax: MiniMax M3 (free)?

Neither wins outright: Google: Gemini 2.5 Flash Lite Preview 09-2025 is the pick when you need audio input, MiniMax: MiniMax M3 (free) when cost matters at scale. In short: Google: Gemini 2.5 Flash Lite Preview 09-2025 when you need audio input, you need file / PDF input; MiniMax: MiniMax M3 (free) when cost matters at scale, you generate long outputs in one go, you want to prototype at zero cost, you want the newer model.

Which is cheaper, Google: Gemini 2.5 Flash Lite Preview 09-2025 or MiniMax: MiniMax M3 (free)?

MiniMax: MiniMax M3 (free) is cheaper: $0 vs $0.18 per 1M tokens at a typical 3:1 input/output mix, 100% less. For 50,000 support conversations a month that is $0 against $15.50.

How much do Google: Gemini 2.5 Flash Lite Preview 09-2025 and MiniMax: MiniMax M3 (free) cost per million tokens?

Google: Gemini 2.5 Flash Lite Preview 09-2025 costs $0.10 per 1M input tokens and $0.40 per 1M output tokens; MiniMax: MiniMax M3 (free) costs $0 input and $0 output. Output tokens are billed separately, so long answers shift the gap.

Which has the bigger context window, Google: Gemini 2.5 Flash Lite Preview 09-2025 or MiniMax: MiniMax M3 (free)?

Both take 1.0M tokens per request — roughly 1,573 pages of text.

Do Google: Gemini 2.5 Flash Lite Preview 09-2025 and MiniMax: MiniMax M3 (free) support function calling and JSON output?

Tool calling — Google: Gemini 2.5 Flash Lite Preview 09-2025: yes, MiniMax: MiniMax M3 (free): yes. JSON-schema structured output — Google: Gemini 2.5 Flash Lite Preview 09-2025: yes, MiniMax: MiniMax M3 (free): yes.

Can Google: Gemini 2.5 Flash Lite Preview 09-2025 or MiniMax: MiniMax M3 (free) read images?

Yes, both accept image input.

Is Google: Gemini 2.5 Flash Lite Preview 09-2025 or MiniMax: MiniMax M3 (free) free to use?

MiniMax: MiniMax M3 (free) has a free variant (minimax/minimax-m3:free). Free variants usually carry tighter rate limits.

What is the maximum output length of Google: Gemini 2.5 Flash Lite Preview 09-2025 and MiniMax: MiniMax M3 (free)?

Google: Gemini 2.5 Flash Lite Preview 09-2025 can write up to 66K tokens (about 49,200 words) per response; MiniMax: MiniMax M3 (free) up to 944K (about 707,800 words).

When were Google: Gemini 2.5 Flash Lite Preview 09-2025 and MiniMax: MiniMax M3 (free) released?

Google: Gemini 2.5 Flash Lite Preview 09-2025 was listed on 2025-09-25; MiniMax: MiniMax M3 (free) on 2026-05-31.

About the two models

Google

Google: Gemini 2.5 Flash Lite Preview 09-2025

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance across common benchmarks compared to earlier Flash models. By default, "thinking" (i.e. multi-pass reasoning) is disabled to prioritize speed, but developers can enable it via the [Reasoning API parameter](https://openrouter.ai/docs/use-cases/reasoning-tokens) to selectively trade off cost for intelligence.

Full Google: Gemini 2.5 Flash Lite Preview 09-2025 specs →

Compare Google: Gemini 2.5 Flash Lite Preview 09-2025 with other models

Compare MiniMax: MiniMax M3 (free) with other models

Advertisement