Qwen2.5-72B-Instruct
Released
2024-06-07
Last refreshed
2026-06-30
Status
Researched 135d ago
Open sourceCommercial use: permittedCodingRAGLong contextClassificationJSON / Tool use
Qwen2.5-72B-Instruct is worth evaluating for coding, rag, and long context when its provider route and context window match the workload.
Use it for
- Teams evaluating coding, rag, and long context
- Workloads that can use a 128k context window
- Buyers comparing 4 tracked provider routes
Do not use it for
- Vision or document-understanding workloads
Specifications
- Family
- Qwen2.5
- Released
- 2024-06-07
- Context
- 128k
- Parameters
- 72.7B
- Architecture
- Decoder Only
- Specialization
- general
- Openness
- Open source
- License
- Apache 2.0OSI-approvedCommercial use: permitted
- Weights
- Available
- Code
- Unknown
- Training
- Fine-tuned
Created by
Pricing
Output / 1M
$0.280
Input / 1M
$0.280
Cheapest of 7 routes · SiliconFlow
About
Instruction-optimized flagship variant for demanding production applications requiring high-accuracy complex problem-solving across industries.
Top use-case fit: coding, agents, and build tasks
Coding
Q/$ B1 relevant benchmark in the decision map.
RAG
Included by capability and metadata signals in the decision map.
Long context
Included by capability and metadata signals in the decision map.
Provider price ladder
Compare all 7Compare API pricing across 4 providers for input and output tokens, batch, and cached reads when available.
| Provider | Input / 1M | Output / 1M | Route |
|---|---|---|---|
| SiliconFlow | $0.280 | $0.280 | Serverless |
| DeepInfra | $0.360 | $0.400 | Serverless |
| OpenRouter | $0.360 | $0.400 | Serverless |
| Novita AI | $0.380 | $0.400 | Serverless |
Available via routers & gateways(1)
Capabilities
Structured Outputs
Benchmark peer barsfor Coding
HumanEvalRank 26 of 97
Benchmark scores(5)
Scores are benchmark-specific and are direction-aware: the same numeric gap can mean very different outcomes across suites. Use the leaderboard context and this model's provider route to decide whether the winning margin is meaningful for your workload.
| Benchmark | Score | Version | Evaluation | Source |
|---|---|---|---|---|
| Google-Proof Q&A | 38.4 | diamondObserved 2024-09-01 | — | Source |
| HellaSwag | 95.6 | standardObserved 2026-03-06 | — | Source |
| HumanEval | 86.6 | pass@1Observed 2024-09-01 | — | Source |
| Massive Multitask Language Understanding | 88.2 | 5-shotObserved 2026-03-06 | — | Source |
| Chatbot Arena | 1270.0 | —Observed 2026-04-15 | — | Source |
Migration checks
No linked migration route is available for this model yet.
Rankings & picks(1)
Compare Qwen2.5-72B-Instruct with other models
- Qwen2.5-72B-Instruct vs Llama 3.3 70B Instruct (free)1K
- Qwen2.5-72B-Instruct vs Llama 3.3 70B859
- Qwen2.5-72B-Instruct vs Qwen3.6-27B706
- Qwen2.5-72B-Instruct vs DeepSeek R1522
- Qwen2.5-72B-Instruct vs Llama 3 70B Instruct461
- Qwen2.5-72B-Instruct vs DeepSeek V4 Flash289
- Qwen2.5-72B-Instruct vs DeepSeek V4 Pro221
- Qwen2.5-72B-Instruct vs GPT-5.5166
Comparison and alternatives
Browse all comparisons →Qwen2.5-72B-Instruct vs Llama 3.3 70BQwen2.5-72B-Instruct vs Mistral Large 2.1 (2411)Qwen2.5-72B-Instruct vs Llama 3.3 70B Instruct (free)Qwen2.5-72B-Instruct vs Qwen3.6-27BQwen2.5-72B-Instruct vs DeepSeek R1Qwen2.5-72B-Instruct vs Llama 3 70B InstructQwen2.5-72B-Instruct vs DeepSeek V4 FlashQwen2.5-72B-Instruct vs DeepSeek V4 ProQwen2.5-72B-Instruct vs GPT-5.5Qwen2.5-72B-Instruct vs DeepSeek V3Qwen2.5-72B-Instruct vs Gemini 2.5 ProQwen2.5-72B-Instruct vs Claude Sonnet 4.6Qwen2.5-72B-Instruct vs GPT-5 ProQwen2.5-72B-Instruct vs Kimi K2.5Qwen2.5-72B-Instruct vs Grok 4Qwen2.5-72B-Instruct vs Grok-3
Show all 23 popular comparisonssorted by 7-day search impressions
Qwen2.5-72B-Instruct vs Qwen3.6-35B-A3B63Qwen2.5-72B-Instruct vs GLM-554Qwen2.5-72B-Instruct vs GLM-5.150Qwen2.5-72B-Instruct vs GLM-5 Turbo48Qwen2.5-72B-Instruct vs Claude Sonnet 4.542Qwen2.5-72B-Instruct vs Gemini 2.5 Flash39Qwen2.5-72B-Instruct vs Claude Opus 4.636Qwen2.5-72B-Instruct vs Mixtral 8x7B35Qwen2.5-72B-Instruct vs Mistral Nemotron34Qwen2.5-72B-Instruct vs Mistral Large 230Qwen2.5-72B-Instruct vs Qwen3.5-397B-A17B20Qwen2.5-72B-Instruct vs Claude Opus 4.717Qwen2.5-72B-Instruct vs GPT-5.413Qwen2.5-72B-Instruct vs Llama 3.2 1B Instruct12Qwen2.5-72B-Instruct vs o39Qwen2.5-72B-Instruct vs Llama 3 8B Instruct8Qwen2.5-72B-Instruct vs Claude Opus 4.57Qwen2.5-72B-Instruct vs Claude 3.7 Sonnet7Qwen2.5-72B-Instruct vs GPT-5.26Qwen2.5-72B-Instruct vs o3 Mini3Qwen2.5-72B-Instruct vs Mistral Large 2 (2407)3Qwen2.5-72B-Instruct vs Mistral Large2Qwen2.5-72B-Instruct vs Code Davinci 0011
Created by
Pricing
Output / 1M
$0.280
Input / 1M
$0.280
Cheapest of 7 routes · SiliconFlow