Atlas / Models / Qwen / Qwen: Qwen3 30B A3B Thinking 2507

Qwen: Qwen3 30B A3B Thinking 2507✓ Catalog verified

qwen/qwen3-30b-a3b-thinking-2507

Qwen3-30B-A3B-Thinking-2507 is a 30B parameter Mixture-of-Experts reasoning model optimized for complex tasks requiring extended multi-step thinking. The model is designed specifically for “thinking mode,” where internal reasoning traces are separated from final answers. Compared to earlier Qwen3-30B releases, this version improves performance across logical reasoning, mathematics, science, coding, and multilingual benchmarks. It also demonstrates stronger instruction following, tool use, and alignment with human preferences. With higher reasoning efficiency and extended output budgets, it is best suited for advanced research, competitive problem solving, and agentic applications requiring structured long-context reasoning.

Input price
$0.08 /1M
Output price
$0.40 /1M
Context
131K
Modalities
text
Released
Aug 28, 2025
Tool calling
✓ Yes
Atlas signal
91/100
01

Overview

Qwen3-30B-A3B-Thinking-2507 is a 30B parameter Mixture-of-Experts reasoning model optimized for complex tasks requiring extended multi-step thinking. The model is designed specifically for “thinking mode,” where internal reasoning traces are separated from final answers. Compared to earlier Qwen3-30B releases, this version improves performance across logical reasoning, mathematics, science, coding, and multilingual benchmarks. It also demonstrates stronger instruction following, tool use, and alignment with human preferences. With higher reasoning efficiency and extended output budgets, it is best suited for advanced research, competitive problem solving, and agentic applications requiring structured long-context reasoning.

Access: available through the official Qwen API. Context window: 131K.

Pricing and metadata from the ModelsAtlas catalog.Last refreshed Aug 5, 2026
02

Specifications

API identifierqwen/qwen3-30b-a3b-thinking-2507
ProviderQwen
Model typeText LLM
Context window131K catalog
Input modalitiestext
Output modalitiestext
ReleasedAug 28, 2025
TokenizerQwen3
ModeratedNo
Architecture modalitytext->text
Supported parameters
frequency_penaltyinclude_reasoninglogit_biasmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_p
Source: provider documentation + catalog feedMethodology →
03

Pricing

Live pricing components for qwen/qwen3-30b-a3b-thinking-2507 as published in the catalog.

$0.08
Input /1M
$0.40
Output /1M
Per request
Per image
Pricing componentRaw unit priceNormalizedUnit
Prompt tokens$8e-8$0.08 / 1MPer input token
Completion tokens$4e-7$0.40 / 1MPer output token
Request feeN/AN/APer request
Image feeN/AN/APer image unit
Web search feeN/AN/APer search request
Source: catalog pricing feedCompare all pricing →Cheapest models →
04

Cost calculator

Estimate monthly spend from your own token volumes.

Scale / Volume

Estimated Total Cost

$0.12Calculated from current list pricing in the ModelsAtlas catalog.
05

Capabilities

CapabilityStatusWhat it means
Visual UnderstandingNot advertisedImage and document analysis support
Audio ProcessingNot advertisedSpeech and voice aligned flows
Tool Calling✓ SupportedSupports tools / function calling
Self-Hosting✓ SupportedDeploy outside managed APIs
Derived from catalog capability tags and supported parameters
06

Capability signals

Directional signals derived from model metadata and capability tags — not official benchmark submissions.

Qwen: Qwen3 30B A3B Thinking 2507
MMLU Signal (Reasoning)
97signal
Coding Signal (HumanEval proxy)
88signal
Math Signal (GSM8K proxy)
92signal
Science Signal (GPQA proxy)
87signal
Estimated from metadata — not an official benchmark runAll benchmarks →Methodology →
07

Quick start

Call Qwen: Qwen3 30B A3B Thinking 2507 through an OpenAI-compatible client.

from openai import OpenAI

client = OpenAI(
    base_url="https://openrouter.ai/api/v1",
    api_key="YOUR_API_KEY"
)

response = client.chat.completions.create(
    model="qwen/qwen3-30b-a3b-thinking-2507",
    messages=[{"role": "user", "content": "Explain quantum physics."}]
)

print(response.choices[0].message.content)
08

Alternatives

Closest models by context window from a different provider.

09

Sources & attribution

Pricing and metadata are maintained in the ModelsAtlas catalog.

Capability bars use metadata tags — directional estimates onlyHow we source and verify data →
10

Frequently asked questions

How much does Qwen: Qwen3 30B A3B Thinking 2507 cost?

$0.08 per 1M input tokens and $0.40 per 1M output tokens on the official Qwen API.

What is the context window of Qwen: Qwen3 30B A3B Thinking 2507?

Qwen: Qwen3 30B A3B Thinking 2507 supports up to 131K tokens of context.

Does Qwen: Qwen3 30B A3B Thinking 2507 support tool / function calling?

Yes, Qwen: Qwen3 30B A3B Thinking 2507 supports tool / function calling — you can register tools and the model will emit structured tool calls.

How do I access Qwen: Qwen3 30B A3B Thinking 2507?

Use the official Qwen API with the model id `qwen/qwen3-30b-a3b-thinking-2507`. See the Quick start section above for code examples.