LLM Reference

Qwen2.5-Coder-32B

Released
2024-11-12
Last refreshed
2026-06-15
Status
Researched 105d ago
Open sourceCommercial use: permittedCodingRAGAgentsLong contextClassificationJSON / Tool use

Qwen2.5-Coder-32B is worth evaluating for coding, rag, and agents when its provider route and context window match the workload.

Use it for

  • Teams evaluating coding, rag, and agents
  • Workloads that can use a 128k context window
  • Buyers comparing 2 tracked provider routes

Do not use it for

  • Vision or document-understanding workloads
Specifications
Released
2024-11-12
Context
128k
Parameters
32B
Architecture
Decoder Only
Knowledge cutoff
2024-02
Specialization
code
Openness
Open source
License
Apache 2.0OSI-approvedCommercial use: permitted
Weights
Available
Code
Unknown
Training
Fine-tuned
Created by

AI research institute of Alibaba Group.

Hangzhou, Zhejiang, China
Founded 2017
Website
Pricing
Output / 1M
$0.200
Input / 1M
$0.200

Cheapest of 2 routes · DeepInfra

About

32B flagship code specialist matching GPT-4o performance with SOTA multi-language repair (75.2% on MdEval) and 3.7% improvement on repo-wide context benchmarks.

Top use-case fit: coding, agents, and build tasks

Coding

Included by capability and metadata signals in the decision map.

RAG

Included by capability and metadata signals in the decision map.

Agents

Included by capability and metadata signals in the decision map.

Provider price ladder

Compare all 2

Compare API pricing across 2 providers for input and output tokens, batch, and cached reads when available.

ProviderInput / 1MOutput / 1MRoute
DeepInfra$0.200$0.200
Serverless
Fireworks AI$0.900$0.900
Serverless

Available via routers & gateways(1)

Capabilities

Structured OutputsCode Execution

Benchmark peer barsfor Coding

No task-mapped benchmark peers are available for this model yet.

Migration checks

No linked migration route is available for this model yet.