Choose AI21: Jamba Large 1.7 if…
No advantage on the data we track — pick it for provider preference or output quality on your own prompts.
Qwen: Qwen3.8 Flash comes out ahead on the data we track: 8 advantages and none for AI21: Jamba Large 1.7.
| Spec | AI21: Jamba Large 1.7 | Qwen: Qwen3.8 Flash |
|---|---|---|
| Provider | AI21 | Qwen |
| Input price (per 1M tokens) | $2.00 | $0.15 |
| Output price (per 1M tokens) | $8.00 | $0.47 |
| Blended price (3:1 input/output) | $3.50 | $0.23 |
| Cheaper than … of all models we track | 17% | 68% |
| Context window | 256K tokens (~384 pages) | 1M tokens (~1,500 pages) |
| Larger context than … of all models | 45% | 70% |
| Max output per response | 4K tokens (~3,100 words) | 131K tokens (~98,300 words) |
| Input types | text | text, image, video |
| Output types | text | text |
| Free variant | No | No |
| Listed since | 2025-08-08 | 2026-08-26 |
Monthly cost at list price for five common workloads. Token counts per unit are stated so you can scale them to your own traffic. A workload that does not fit a model's context window is marked rather than priced.
| Workload (per month) | AI21: Jamba Large 1.7 | Qwen: Qwen3.8 Flash | Difference |
|---|---|---|---|
| Customer-support chatbot 50,000 conversations — each a 6-turn conversation (1,500 in / 400 out tokens) | $310 | $20.65 | Qwen: Qwen3.8 Flash saves $289 |
| RAG search over documents 20,000 questions — each a question answered from ~8 retrieved passages (6,000 in / 500 out tokens) | $320 | $22.70 | Qwen: Qwen3.8 Flash saves $297 |
| Coding / tool-using agent 2,000 tasks — each a multi-step task with tool calls and file context (60,000 in / 6,000 out tokens) | $336 | $23.64 | Qwen: Qwen3.8 Flash saves $312 |
| Batch summarisation 10,000 documents — each a 6,000-word report summarised to a page (8,000 in / 600 out tokens) | $208 | $14.82 | Qwen: Qwen3.8 Flash saves $193 |
| Long-document analysis 500 documents — each a 150-page contract or codebase read in one pass (100,000 in / 2,000 out tokens) | $108 | $7.97 | Qwen: Qwen3.8 Flash saves $100 |
| Capability | AI21: Jamba Large 1.7 | Qwen: Qwen3.8 Flash | What it lets you do |
|---|---|---|---|
| Tool / function calling | ✓ Yes | ✓ Yes | call your APIs and run agent loops |
| Structured output (JSON schema) | ✓ Yes | ✓ Yes | return JSON that validates against your schema |
| Reasoning / thinking mode | — | ✓ Yes | spend extra tokens thinking before answering hard problems |
| Image input | — | ✓ Yes | read screenshots, charts and scanned pages |
| Audio input | — | — | take speech or audio directly |
| File / PDF input | — | — | accept documents without your own parsing step |
| Image output | — | — | generate images, not just text |
| Built-in web search | — | — | answer from live web results |
| Seed (reproducible sampling) | — | ✓ Yes | repeat a generation for tests and evals |
| Log-probabilities | — | ✓ Yes | read token confidence for classification and scoring |
From each model's published parameters and input/output types. “—” means not advertised, not proven absent.
| Date | Model | Input / 1M | Output / 1M |
|---|---|---|---|
| 2026-08-28 | Qwen: Qwen3.8 Flash | $0.15 | $0.47 |
No advantage on the data we track — pick it for provider preference or output quality on your own prompts.
Qwen: Qwen3.8 Flash comes out ahead on the data we track: 8 advantages and none for AI21: Jamba Large 1.7.
Qwen: Qwen3.8 Flash is cheaper: $0.23 vs $3.50 per 1M tokens at a typical 3:1 input/output mix, 93% less. For 50,000 support conversations a month that is $20.65 against $310.
AI21: Jamba Large 1.7 costs $2.00 per 1M input tokens and $8.00 per 1M output tokens; Qwen: Qwen3.8 Flash costs $0.15 input and $0.47 output. Output tokens are billed separately, so long answers shift the gap.
Qwen: Qwen3.8 Flash: 1M tokens, about 1,500 pages, against 256K (about 384 pages) for AI21: Jamba Large 1.7.
Tool calling — AI21: Jamba Large 1.7: yes, Qwen: Qwen3.8 Flash: yes. JSON-schema structured output — AI21: Jamba Large 1.7: yes, Qwen: Qwen3.8 Flash: yes.
Only Qwen: Qwen3.8 Flash accepts image input; AI21: Jamba Large 1.7 is text-only for input.
No free variant is listed for either; both are pay-per-token.
AI21: Jamba Large 1.7 can write up to 4K tokens (about 3,100 words) per response; Qwen: Qwen3.8 Flash up to 131K (about 98,300 words).
AI21: Jamba Large 1.7 was listed on 2025-08-08; Qwen: Qwen3.8 Flash on 2026-08-26.
AI21: Jamba Large 1.7: price unchanged since we started tracking it on 2026-03-10. Qwen: Qwen3.8 Flash: input price cut 6% ($0.16 → $0.15/1M) on 2026-08-28.
Jamba Large 1.7 is the latest model in the Jamba open family, offering improvements in grounding, instruction-following, and overall efficiency. Built on a hybrid SSM-Transformer architecture with a 256K context...
Full AI21: Jamba Large 1.7 specs →Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.
Full Qwen: Qwen3.8 Flash specs →