Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...
Model details →Inception: Mercury 2.5 Preview vs Qwen: Qwen3 VL 30B A3B Thinking
Inception: Mercury 2.5 Preview and Qwen: Qwen3 VL 30B A3B Thinking are evenly matched on 6 axes — pick the one whose provider, latency, or licensing fits your stack.
Qwen3-VL-30B-A3B-Thinking is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Thinking variant enhances reasoning in STEM, math, and complex tasks. It excels...
Model details →Side-by-side comparison
| Capability | Inception: Mercury 2.5 Preview | Qwen: Qwen3 VL 30B A3B Thinking | Winner |
|---|---|---|---|
| Context window Maximum number of input tokens the model can attend to in a single request. | 260K | 262K | 🏆 Qwen: Qwen3 VL 30B A3B Thinking |
| Input price (per 1M) Cost per million input tokens billed by the provider. | $0.04 | $0.20 | 🏆 Inception: Mercury 2.5 Preview |
| Output price (per 1M) Cost per million output tokens billed by the provider. | $0.15 | $2.40 | 🏆 Inception: Mercury 2.5 Preview |
| Tool / function calling First-class support for emitting structured tool calls. | Yes | Yes | Tie |
| Vision input Accepts image inputs alongside text. | — | Yes | 🏆 Qwen: Qwen3 VL 30B A3B Thinking |
| Reasoning mode Internal chain-of-thought / extended-thinking support. | Yes | Yes | Tie |
Frequently asked questions
Is Inception: Mercury 2.5 Preview better than Qwen: Qwen3 VL 30B A3B Thinking?
Inception: Mercury 2.5 Preview and Qwen: Qwen3 VL 30B A3B Thinking are evenly matched on 6 axes — pick the one whose provider, latency, or licensing fits your stack.
What's the price difference between Inception: Mercury 2.5 Preview and Qwen: Qwen3 VL 30B A3B Thinking?
Input: $0.04 vs $0.20 per 1M tokens. Output: $0.15 vs $2.40 per 1M tokens.
What context windows do Inception: Mercury 2.5 Preview and Qwen: Qwen3 VL 30B A3B Thinking support?
Inception: Mercury 2.5 Preview supports up to 260K tokens. Qwen: Qwen3 VL 30B A3B Thinking supports up to 262K tokens.
Do both Inception: Mercury 2.5 Preview and Qwen: Qwen3 VL 30B A3B Thinking support tool calling?
Inception: Mercury 2.5 Preview: yes. Qwen: Qwen3 VL 30B A3B Thinking: yes.