ERNIE-4.5-VL-424B-A47B is a multimodal Mixture-of-Experts (MoE) model from Baidu’s ERNIE 4.5 series, featuring 424B total parameters with 47B active per token. It is trained jointly on text and image data using a heterogeneous MoE architect…
Model details →Baidu: ERNIE 4.5 VL 424B A47B vs Qwen: Qwen3 VL 30B A3B Thinking
Qwen: Qwen3 VL 30B A3B Thinking wins on 3 of 6 axes — pricing and capability skew in its favour for most workloads.
Qwen3-VL-30B-A3B-Thinking is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Thinking variant enhances reasoning in STEM, math, and complex tasks. It excels in perception of real-w…
Model details →Side-by-side comparison
| Capability | Baidu: ERNIE 4.5 VL 424B A47B | Qwen: Qwen3 VL 30B A3B Thinking | Winner |
|---|---|---|---|
| Context window Maximum number of input tokens the model can attend to in a single request. | 123K | 131K | 🏆 Qwen: Qwen3 VL 30B A3B Thinking |
| Input price (per 1M) Cost per million input tokens billed by the provider. | $0.42 | $0.13 | 🏆 Qwen: Qwen3 VL 30B A3B Thinking |
| Output price (per 1M) Cost per million output tokens billed by the provider. | $1.25 | $1.56 | 🏆 Baidu: ERNIE 4.5 VL 424B A47B |
| Tool / function calling First-class support for emitting structured tool calls. | — | Yes | 🏆 Qwen: Qwen3 VL 30B A3B Thinking |
| Vision input Accepts image inputs alongside text. | Yes | Yes | Tie |
| Reasoning mode Internal chain-of-thought / extended-thinking support. | Yes | Yes | Tie |
Frequently asked questions
Is Baidu: ERNIE 4.5 VL 424B A47B better than Qwen: Qwen3 VL 30B A3B Thinking?
Qwen: Qwen3 VL 30B A3B Thinking wins on 3 of 6 axes — pricing and capability skew in its favour for most workloads.
What's the price difference between Baidu: ERNIE 4.5 VL 424B A47B and Qwen: Qwen3 VL 30B A3B Thinking?
Input: $0.42 vs $0.13 per 1M tokens. Output: $1.25 vs $1.56 per 1M tokens.
What context windows do Baidu: ERNIE 4.5 VL 424B A47B and Qwen: Qwen3 VL 30B A3B Thinking support?
Baidu: ERNIE 4.5 VL 424B A47B supports up to 123K tokens. Qwen: Qwen3 VL 30B A3B Thinking supports up to 131K tokens.
Do both Baidu: ERNIE 4.5 VL 424B A47B and Qwen: Qwen3 VL 30B A3B Thinking support tool calling?
Baidu: ERNIE 4.5 VL 424B A47B: not advertised. Qwen: Qwen3 VL 30B A3B Thinking: yes.