QwQ 32B
QwQ 32B is worth evaluating for long context when its provider route and context window match the workload.
Use it for
- Teams evaluating long context
- Workloads that can use a 128k context window
- Buyers comparing 2 tracked provider routes
Do not use it for
- Vision or document-understanding workloads
- Strict JSON or tool-calling flows
- Family
- QwQ
- Released
- 2025-03-05
- Context
- 128k
- Parameters
- 32.5B
- Architecture
- Decoder Only
- Specialization
- reasoning
- Training
- pretrained
Cheapest of 2 routes · Cloudflare Workers AI
About
QwQ-32B is the first full release in Alibaba's QwQ reasoning series. Built on a 32.5B-parameter dense transformer, it achieves significantly enhanced performance on complex tasks—mathematics, coding, and multi-step reasoning—through extended chain-of-thought thinking. Available open-weight on Hugging Face, it delivers frontier reasoning in an efficient package.
QwQ 32B is an open-source model in the QwQ family. The structured metadata tracks a 128k-token context window and reasoning. This page tracks provider routes through Cloudflare Workers AI and OpenRouter, with the cheapest tracked route listed at $0.66 input and $1 output per 1M tokens. No headline benchmark score is tracked for QwQ 32B yet.
Top use-case fit
Long context
Included by capability and metadata signals in the decision map.
Provider price ladder
Compare all 2Compare API pricing across 2 providers for input and output tokens, batch, and cached reads when available.
| Provider | Input / 1M | Output / 1M | Route |
|---|---|---|---|
| Cloudflare Workers AI | $0.660 | $1.00 | Serverless |
| OpenRouter | - | - | ServerlessPartial |
Capabilities
Benchmark peer barsfor Long context
No task-mapped benchmark peers are available for this model yet.
Migration checks
No linked migration route is available for this model yet.