Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...
Model details →Google: Gemma 4 26B A4B vs Inception: Mercury 2.5 Preview
Google: Gemma 4 26B A4B and Inception: Mercury 2.5 Preview are evenly matched on 6 axes — pick the one whose provider, latency, or licensing fits your stack.
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...
Model details →Side-by-side comparison
| Capability | Google: Gemma 4 26B A4B | Inception: Mercury 2.5 Preview | Winner |
|---|---|---|---|
| Context window Maximum number of input tokens the model can attend to in a single request. | 262K | 260K | 🏆 Google: Gemma 4 26B A4B |
| Input price (per 1M) Cost per million input tokens billed by the provider. | $0.07 | $0.04 | 🏆 Inception: Mercury 2.5 Preview |
| Output price (per 1M) Cost per million output tokens billed by the provider. | $0.34 | $0.15 | 🏆 Inception: Mercury 2.5 Preview |
| Tool / function calling First-class support for emitting structured tool calls. | Yes | Yes | Tie |
| Vision input Accepts image inputs alongside text. | Yes | — | 🏆 Google: Gemma 4 26B A4B |
| Reasoning mode Internal chain-of-thought / extended-thinking support. | Yes | Yes | Tie |
Frequently asked questions
Is Google: Gemma 4 26B A4B better than Inception: Mercury 2.5 Preview?
Google: Gemma 4 26B A4B and Inception: Mercury 2.5 Preview are evenly matched on 6 axes — pick the one whose provider, latency, or licensing fits your stack.
What's the price difference between Google: Gemma 4 26B A4B and Inception: Mercury 2.5 Preview?
Input: $0.07 vs $0.04 per 1M tokens. Output: $0.34 vs $0.15 per 1M tokens.
What context windows do Google: Gemma 4 26B A4B and Inception: Mercury 2.5 Preview support?
Google: Gemma 4 26B A4B supports up to 262K tokens. Inception: Mercury 2.5 Preview supports up to 260K tokens.
Do both Google: Gemma 4 26B A4B and Inception: Mercury 2.5 Preview support tool calling?
Google: Gemma 4 26B A4B: yes. Inception: Mercury 2.5 Preview: yes.