Google: Gemini 2.5 Flash Lite Preview 09-2025 vs Qwen: Qwen3.8 Flash

Neither wins outright: Google: Gemini 2.5 Flash Lite Preview 09-2025 is the pick when cost matters at scale, Qwen: Qwen3.8 Flash when you need log-probabilities.

Key numbers

SpecGoogle: Gemini 2.5 Flash Lite Preview 09-2025Qwen: Qwen3.8 Flash
ProviderGoogleQwen
Input price (per 1M tokens)$0.10$0.15
Output price (per 1M tokens)$0.40$0.47
Blended price (3:1 input/output)$0.18$0.23
Cheaper than … of all models we track72%68%
Context window1.0M tokens (~1,573 pages)1M tokens (~1,500 pages)
Larger context than … of all models81%70%
Max output per response66K tokens (~49,200 words)131K tokens (~98,300 words)
Input typestext, image, file, audio, videotext, image, video
Output typestexttext
Free variantNoNo
Listed since2025-09-252026-08-26

What it costs for real workloads

Monthly cost at list price for five common workloads. Token counts per unit are stated so you can scale them to your own traffic. A workload that does not fit a model's context window is marked rather than priced.

Workload (per month)Google: Gemini 2.5 Flash Lite Preview 09-2025Qwen: Qwen3.8 FlashDifference
Customer-support chatbot
50,000 conversations — each a 6-turn conversation (1,500 in / 400 out tokens)
$15.50$20.65Google: Gemini 2.5 Flash Lite Preview 09-2025 saves $5.15
RAG search over documents
20,000 questions — each a question answered from ~8 retrieved passages (6,000 in / 500 out tokens)
$16.00$22.70Google: Gemini 2.5 Flash Lite Preview 09-2025 saves $6.70
Coding / tool-using agent
2,000 tasks — each a multi-step task with tool calls and file context (60,000 in / 6,000 out tokens)
$16.80$23.64Google: Gemini 2.5 Flash Lite Preview 09-2025 saves $6.84
Batch summarisation
10,000 documents — each a 6,000-word report summarised to a page (8,000 in / 600 out tokens)
$10.40$14.82Google: Gemini 2.5 Flash Lite Preview 09-2025 saves $4.42
Long-document analysis
500 documents — each a 150-page contract or codebase read in one pass (100,000 in / 2,000 out tokens)
$5.40$7.97Google: Gemini 2.5 Flash Lite Preview 09-2025 saves $2.57

Capabilities side by side

CapabilityGoogle: Gemini 2.5 Flash Lite Preview 09-2025Qwen: Qwen3.8 FlashWhat it lets you do
Tool / function calling✓ Yes✓ Yescall your APIs and run agent loops
Structured output (JSON schema)✓ Yes✓ Yesreturn JSON that validates against your schema
Reasoning / thinking mode✓ Yes✓ Yesspend extra tokens thinking before answering hard problems
Image input✓ Yes✓ Yesread screenshots, charts and scanned pages
Audio input✓ Yes—take speech or audio directly
File / PDF input✓ Yes—accept documents without your own parsing step
Image output——generate images, not just text
Built-in web search——answer from live web results
Seed (reproducible sampling)✓ Yes✓ Yesrepeat a generation for tests and evals
Log-probabilities—✓ Yesread token confidence for classification and scoring

From each model's published parameters and input/output types. “—” means not advertised, not proven absent.

Price history

DateModelInput / 1MOutput / 1M
2026-08-28Qwen: Qwen3.8 Flash$0.15$0.47

Which should you choose?

Choose Google: Gemini 2.5 Flash Lite Preview 09-2025 if…

  • cost matters at scale — it is 24% cheaper on a typical 3:1 input/output mix ($15.50 vs $20.65 a month for 50,000 support chats)
  • you need audio input — to take speech or audio directly; Qwen: Qwen3.8 Flash does not advertise it
  • you need file / PDF input — to accept documents without your own parsing step; Qwen: Qwen3.8 Flash does not advertise it

Choose Qwen: Qwen3.8 Flash if…

  • you need log-probabilities — to read token confidence for classification and scoring; Google: Gemini 2.5 Flash Lite Preview 09-2025 does not advertise it
  • you generate long outputs in one go — up to 131K tokens (about 98,300 words) per response
  • you want the newer model — released 2026-08-26, 11 months after the other

Frequently asked questions

Is Google: Gemini 2.5 Flash Lite Preview 09-2025 better than Qwen: Qwen3.8 Flash?

Neither wins outright: Google: Gemini 2.5 Flash Lite Preview 09-2025 is the pick when cost matters at scale, Qwen: Qwen3.8 Flash when you need log-probabilities. In short: Google: Gemini 2.5 Flash Lite Preview 09-2025 when cost matters at scale, you need audio input, you need file / PDF input; Qwen: Qwen3.8 Flash when you need log-probabilities, you generate long outputs in one go, you want the newer model.

Which is cheaper, Google: Gemini 2.5 Flash Lite Preview 09-2025 or Qwen: Qwen3.8 Flash?

Google: Gemini 2.5 Flash Lite Preview 09-2025 is cheaper: $0.18 vs $0.23 per 1M tokens at a typical 3:1 input/output mix, 24% less. For 50,000 support conversations a month that is $15.50 against $20.65.

How much do Google: Gemini 2.5 Flash Lite Preview 09-2025 and Qwen: Qwen3.8 Flash cost per million tokens?

Google: Gemini 2.5 Flash Lite Preview 09-2025 costs $0.10 per 1M input tokens and $0.40 per 1M output tokens; Qwen: Qwen3.8 Flash costs $0.15 input and $0.47 output. Output tokens are billed separately, so long answers shift the gap.

Which has the bigger context window, Google: Gemini 2.5 Flash Lite Preview 09-2025 or Qwen: Qwen3.8 Flash?

Google: Gemini 2.5 Flash Lite Preview 09-2025: 1.0M tokens, about 1,573 pages, against 1M (about 1,500 pages) for Qwen: Qwen3.8 Flash.

Do Google: Gemini 2.5 Flash Lite Preview 09-2025 and Qwen: Qwen3.8 Flash support function calling and JSON output?

Tool calling — Google: Gemini 2.5 Flash Lite Preview 09-2025: yes, Qwen: Qwen3.8 Flash: yes. JSON-schema structured output — Google: Gemini 2.5 Flash Lite Preview 09-2025: yes, Qwen: Qwen3.8 Flash: yes.

Can Google: Gemini 2.5 Flash Lite Preview 09-2025 or Qwen: Qwen3.8 Flash read images?

Yes, both accept image input.

Is Google: Gemini 2.5 Flash Lite Preview 09-2025 or Qwen: Qwen3.8 Flash free to use?

No free variant is listed for either; both are pay-per-token.

What is the maximum output length of Google: Gemini 2.5 Flash Lite Preview 09-2025 and Qwen: Qwen3.8 Flash?

Google: Gemini 2.5 Flash Lite Preview 09-2025 can write up to 66K tokens (about 49,200 words) per response; Qwen: Qwen3.8 Flash up to 131K (about 98,300 words).

When were Google: Gemini 2.5 Flash Lite Preview 09-2025 and Qwen: Qwen3.8 Flash released?

Google: Gemini 2.5 Flash Lite Preview 09-2025 was listed on 2025-09-25; Qwen: Qwen3.8 Flash on 2026-08-26.

Has the price of Google: Gemini 2.5 Flash Lite Preview 09-2025 or Qwen: Qwen3.8 Flash changed?

Google: Gemini 2.5 Flash Lite Preview 09-2025: price unchanged since we started tracking it on 2026-03-10. Qwen: Qwen3.8 Flash: input price cut 6% ($0.16 → $0.15/1M) on 2026-08-28.

About the two models

Google

Google: Gemini 2.5 Flash Lite Preview 09-2025

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance across common benchmarks compared to earlier Flash models. By default, "thinking" (i.e. multi-pass reasoning) is disabled to prioritize speed, but developers can enable it via the [Reasoning API parameter](https://openrouter.ai/docs/use-cases/reasoning-tokens) to selectively trade off cost for intelligence.

Full Google: Gemini 2.5 Flash Lite Preview 09-2025 specs →
Qwen

Qwen: Qwen3.8 Flash

Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.

Full Qwen: Qwen3.8 Flash specs →

Compare Google: Gemini 2.5 Flash Lite Preview 09-2025 with other models

Compare Qwen: Qwen3.8 Flash with other models

Advertisement