Atlas / Models / Qwen / Qwen2.5 72B Instruct

Qwen2.5 72B Instruct✓ Catalog verified

qwen/qwen-2.5-72b-instruct

Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and mathematics, thanks to our specialized expert models in these domains. - Significant improvements in instruction following, generating long texts (over 8K tokens), understanding structured data (e.g, tables), and generating structured outputs especially JSON. More resilient to the diversity of system prompts, enhancing role-play implementation and condition-setting for chatbots. - Long-context Support up to 128K tokens and can generate up to 8K tokens. - Multilingual support for over 29 languages, including Chinese, English, French, Spanish, Portuguese, German, Italian, Russian, Japanese, Korean, Vietnamese, Thai, Arabic, and more. Usage of this model is subject to Tongyi Qianwen LICENSE AGREEMENT.

Input price
$0.12 /1M
Output price
$0.39 /1M
Context
33K
Modalities
text
Released
Sep 19, 2024
Tool calling
✓ Yes
Atlas signal
79/100
01

Overview

Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and mathematics, thanks to our specialized expert models in these domains. - Significant improvements in instruction following, generating long texts (over 8K tokens), understanding structured data (e.g, tables), and generating structured outputs especially JSON. More resilient to the diversity of system prompts, enhancing role-play implementation and condition-setting for chatbots. - Long-context Support up to 128K tokens and can generate up to 8K tokens. - Multilingual support for over 29 languages, including Chinese, English, French, Spanish, Portuguese, German, Italian, Russian, Japanese, Korean, Vietnamese, Thai, Arabic, and more. Usage of this model is subject to Tongyi Qianwen LICENSE AGREEMENT.

Access: available through the official Qwen API. Context window: 33K.

Pricing and metadata from the ModelsAtlas catalog.Last refreshed Aug 5, 2026
02

Specifications

API identifierqwen/qwen-2.5-72b-instruct
ProviderQwen
Model typeText LLM
Context window33K catalog
Input modalitiestext
Output modalitiestext
ReleasedSep 19, 2024
TokenizerQwen
Instruction formatchatml
ModeratedNo
Architecture modalitytext->text
Supported parameters
frequency_penaltymax_tokensmin_ppresence_penaltyrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_p
Source: provider documentation + catalog feedMethodology →
03

Pricing

Live pricing components for qwen/qwen-2.5-72b-instruct as published in the catalog.

$0.12
Input /1M
$0.39
Output /1M
Per request
Per image
Pricing componentRaw unit priceNormalizedUnit
Prompt tokens$1.2e-7$0.12 / 1MPer input token
Completion tokens$3.9e-7$0.39 / 1MPer output token
Request feeN/AN/APer request
Image feeN/AN/APer image unit
Web search feeN/AN/APer search request
Source: catalog pricing feedCompare all pricing →Cheapest models →
04

Cost calculator

Estimate monthly spend from your own token volumes.

Scale / Volume

Estimated Total Cost

$0.14Calculated from current list pricing in the ModelsAtlas catalog.
05

Capabilities

CapabilityStatusWhat it means
Visual UnderstandingNot advertisedImage and document analysis support
Audio ProcessingNot advertisedSpeech and voice aligned flows
Tool Calling✓ SupportedSupports tools / function calling
Self-Hosting✓ SupportedDeploy outside managed APIs
Derived from catalog capability tags and supported parameters
06

Capability signals

Directional signals derived from model metadata and capability tags — not official benchmark submissions.

Qwen2.5 72B Instruct
MMLU Signal (Reasoning)
74signal
Coding Signal (HumanEval proxy)
83signal
Math Signal (GSM8K proxy)
83signal
Science Signal (GPQA proxy)
74signal
Estimated from metadata — not an official benchmark runAll benchmarks →Methodology →
07

Quick start

Call Qwen2.5 72B Instruct through an OpenAI-compatible client.

from openai import OpenAI

client = OpenAI(
    base_url="https://openrouter.ai/api/v1",
    api_key="YOUR_API_KEY"
)

response = client.chat.completions.create(
    model="qwen/qwen-2.5-72b-instruct",
    messages=[{"role": "user", "content": "Explain quantum physics."}]
)

print(response.choices[0].message.content)
08

Alternatives

Closest models by context window from a different provider.

09

Sources & attribution

Pricing and metadata are maintained in the ModelsAtlas catalog.

Capability bars use metadata tags — directional estimates onlyHow we source and verify data →
10

Frequently asked questions

How much does Qwen2.5 72B Instruct cost?

$0.12 per 1M input tokens and $0.39 per 1M output tokens on the official Qwen API.

What is the context window of Qwen2.5 72B Instruct?

Qwen2.5 72B Instruct supports up to 33K tokens of context.

Does Qwen2.5 72B Instruct support tool / function calling?

Yes, Qwen2.5 72B Instruct supports tool / function calling — you can register tools and the model will emit structured tool calls.

How do I access Qwen2.5 72B Instruct?

Use the official Qwen API with the model id `qwen/qwen-2.5-72b-instruct`. See the Quick start section above for code examples.