Qwen2.5-72B
Released
2025-10-10
Last refreshed
2026-09-22
Status
Researched 10d ago
Open sourceCommercial use: permittedRAGAgentsLong contextClassificationJSON / Tool use
Qwen2.5-72B is a released rag, agents, and long context model with open-source and 128k context; evaluate it while provider pricing coverage matures.
Alibaba releases · 46 in the last 12 months · this family litChangelog →
Specifications
- Family
- Qwen2.5
- Released
- 2025-10-10
- Context
- 128k
- Parameters
- 72B
- Knowledge cutoff
- 2024-09
- Openness
- Open source
- License
- Apache 2.0OSI-approvedCommercial use: permitted
- Weights
- Available
- Code
- Unknown
Created by
Pricing
No tracked provider token pricing is available yet.
About
Alibaba Qwen2.5 72B model with improved reasoning and 128K context window. Open-weights, optimized for diverse tasks including coding and analysis.
Provider price ladder
No tracked provider token pricing is available for this model yet.
Capabilities
JSON / Tool useFine-tuning
Benchmark peer barsfor Classification
MMLU PRORank 69 of 105
Benchmark scores(1)
Scores are benchmark-specific and are direction-aware: the same numeric gap can mean very different outcomes across suites. Use the leaderboard context and this model's provider route to decide whether the winning margin is meaningful for your workload.
Migration checks
No linked migration route is available for this model yet.
Compare Qwen2.5-72B with other models
Comparison and alternatives
Browse all comparisons →Qwen2.5-72B vs Llama 3.3 70BQwen2.5-72B vs DeepSeek V3Qwen2.5-72B vs Llama 3.1 405BQwen2.5-72B vs DeepSeek V4 FlashQwen2.5-72B vs Qwen3.6-27BQwen2.5-72B vs DeepSeek R1Qwen2.5-72B vs Grok 4Qwen2.5-72B vs DeepSeek V3.1Qwen2.5-72B vs Llama 3.1 405B InstructQwen2.5-72B vs Together AI - Llama 3 8B LiteQwen2.5-72B vs GLM-5.1Qwen2.5-72B vs Claude Opus 4.7Qwen2.5-72B vs DeepSeek V4 ProQwen2.5-72B vs Claude Sonnet 4.6Qwen2.5-72B vs GPT-5.5Qwen2.5-72B vs Llama 3 70B Instruct
Show all 22 popular comparisonssorted by 7-day search impressions
Qwen2.5-72B vs Llama 3 8B Instruct40Qwen2.5-72B vs GPT-5.5 Instant39Qwen2.5-72B vs Claude Sonnet 4.533Qwen2.5-72B vs Claude Opus 4.626Qwen2.5-72B vs Grok Build 0.119Qwen2.5-72B vs Claude Opus 4.514Qwen2.5-72B vs GLM-5 9B14Qwen2.5-72B vs GLM-5 Turbo12Qwen2.5-72B vs Code Cushman 00212Qwen2.5-72B vs Mistral Large 212Qwen2.5-72B vs Claude Haiku 4.512Qwen2.5-72B vs o3 Deep Research8Qwen2.5-72B vs Grok 4.38Qwen2.5-72B vs Kimi K2.58Qwen2.5-72B vs GPT-5.27Qwen2.5-72B vs GPT-5.47Qwen2.5-72B vs GLM-5V-Turbo7Qwen2.5-72B vs GPT-5.4-Cyber5Qwen2.5-72B vs Gemini 2.5 Flash Live API4Qwen2.5-72B vs Mistral Large 2 (2407)3Qwen2.5-72B vs Mistral Large 3 675B Instruct3Qwen2.5-72B vs Code Davinci 0012