Choose DeepSeek: DeepSeek V4 Flash Vision Exp (batch) if…
- you need log-probabilities — to read token confidence for classification and scoring; MiniMax: MiniMax M3 (free) does not advertise it
Neither wins outright: DeepSeek: DeepSeek V4 Flash Vision Exp (batch) is the pick when you need log-probabilities, MiniMax: MiniMax M3 (free) when cost matters at scale.
| Spec | DeepSeek: DeepSeek V4 Flash Vision Exp (batch) | MiniMax: MiniMax M3 (free) |
|---|---|---|
| Provider | DeepSeek | Minimax |
| Input price (per 1M tokens) | $0.11 | $0 |
| Output price (per 1M tokens) | $0.33 | $0 |
| Blended price (3:1 input/output) | $0.17 | $0 |
| Cheaper than … of all models we track | 73% | 90% |
| Context window | 1.0M tokens (~1,573 pages) | 1.0M tokens (~1,573 pages) |
| Larger context than … of all models | 81% | 81% |
| Max output per response | 944K tokens (~707,800 words) | 944K tokens (~707,800 words) |
| Input types | text, image | text, image, video |
| Output types | text | text |
| Free variant | No | Yes |
| Listed since | 2026-08-21 | 2026-05-31 |
Monthly cost at list price for five common workloads. Token counts per unit are stated so you can scale them to your own traffic. A workload that does not fit a model's context window is marked rather than priced.
| Workload (per month) | DeepSeek: DeepSeek V4 Flash Vision Exp (batch) | MiniMax: MiniMax M3 (free) | Difference |
|---|---|---|---|
| Customer-support chatbot 50,000 conversations — each a 6-turn conversation (1,500 in / 400 out tokens) | $14.85 | $0 | MiniMax: MiniMax M3 (free) saves $14.85 |
| RAG search over documents 20,000 questions — each a question answered from ~8 retrieved passages (6,000 in / 500 out tokens) | $16.50 | $0 | MiniMax: MiniMax M3 (free) saves $16.50 |
| Coding / tool-using agent 2,000 tasks — each a multi-step task with tool calls and file context (60,000 in / 6,000 out tokens) | $17.16 | $0 | MiniMax: MiniMax M3 (free) saves $17.16 |
| Batch summarisation 10,000 documents — each a 6,000-word report summarised to a page (8,000 in / 600 out tokens) | $10.78 | $0 | MiniMax: MiniMax M3 (free) saves $10.78 |
| Long-document analysis 500 documents — each a 150-page contract or codebase read in one pass (100,000 in / 2,000 out tokens) | $5.83 | $0 | MiniMax: MiniMax M3 (free) saves $5.83 |
| Capability | DeepSeek: DeepSeek V4 Flash Vision Exp (batch) | MiniMax: MiniMax M3 (free) | What it lets you do |
|---|---|---|---|
| Tool / function calling | ✓ Yes | ✓ Yes | call your APIs and run agent loops |
| Structured output (JSON schema) | ✓ Yes | ✓ Yes | return JSON that validates against your schema |
| Reasoning / thinking mode | ✓ Yes | ✓ Yes | spend extra tokens thinking before answering hard problems |
| Image input | ✓ Yes | ✓ Yes | read screenshots, charts and scanned pages |
| Audio input | — | — | take speech or audio directly |
| File / PDF input | — | — | accept documents without your own parsing step |
| Image output | — | — | generate images, not just text |
| Built-in web search | — | — | answer from live web results |
| Seed (reproducible sampling) | — | ✓ Yes | repeat a generation for tests and evals |
| Log-probabilities | ✓ Yes | — | read token confidence for classification and scoring |
From each model's published parameters and input/output types. “—” means not advertised, not proven absent.
Neither wins outright: DeepSeek: DeepSeek V4 Flash Vision Exp (batch) is the pick when you need log-probabilities, MiniMax: MiniMax M3 (free) when cost matters at scale. In short: DeepSeek: DeepSeek V4 Flash Vision Exp (batch) when you need log-probabilities; MiniMax: MiniMax M3 (free) when cost matters at scale, you need seed (reproducible sampling), you want to prototype at zero cost.
MiniMax: MiniMax M3 (free) is cheaper: $0 vs $0.17 per 1M tokens at a typical 3:1 input/output mix, 100% less. For 50,000 support conversations a month that is $0 against $14.85.
DeepSeek: DeepSeek V4 Flash Vision Exp (batch) costs $0.11 per 1M input tokens and $0.33 per 1M output tokens; MiniMax: MiniMax M3 (free) costs $0 input and $0 output. Output tokens are billed separately, so long answers shift the gap.
Both take 1.0M tokens per request — roughly 1,573 pages of text.
Tool calling — DeepSeek: DeepSeek V4 Flash Vision Exp (batch): yes, MiniMax: MiniMax M3 (free): yes. JSON-schema structured output — DeepSeek: DeepSeek V4 Flash Vision Exp (batch): yes, MiniMax: MiniMax M3 (free): yes.
Yes, both accept image input.
MiniMax: MiniMax M3 (free) has a free variant (minimax/minimax-m3:free). Free variants usually carry tighter rate limits.
DeepSeek: DeepSeek V4 Flash Vision Exp (batch) can write up to 944K tokens (about 707,800 words) per response; MiniMax: MiniMax M3 (free) up to 944K (about 707,800 words).
DeepSeek: DeepSeek V4 Flash Vision Exp (batch) was listed on 2026-08-21; MiniMax: MiniMax M3 (free) on 2026-05-31.
DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,...
Full DeepSeek: DeepSeek V4 Flash Vision Exp (batch) specs →MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
Full MiniMax: MiniMax M3 (free) specs →