Last refreshed 2026-06-15. Next refresh: weekly.
Why use QwQ 32B on OpenRouter?
OpenRouter offers QwQ 32B with competitive pricing. OpenRouter is a multi-provider LLM aggregator offering unified API access to 300+ models from all major labs and emerging providers, with automatic failover for reliability.
Compare QwQ 32B across 2 providers to find the best fit for your use caseInput / 1M
-
Output / 1M
-
Cache
Not sourced
Batch
Not sourced
Setup recipe
Docs fallbackInstall
Use the provider REST API or SDKAuth
Create a provider API keyCall
model: qwen/qwq-32bModel ID
qwen/qwq-32bRequest example
Curated snippets for this provider are not sourced yet. Use OpenRouter documentation with model ID
qwen/qwq-32b.Gotchas
- Use provider model ID "qwen/qwq-32b", not the LLMReference slug "qwq-32b".
Compare QwQ 32B Across Providers
| Provider | Input (per 1M) | Output (per 1M) |
|---|---|---|
| Cloudflare Workers AI | $0.66 | $1.00 |
| OpenRouter | — | — |
Capabilities
Reasoning
About QwQ 32B
QwQ-32B is the first full release in Alibaba's QwQ reasoning series. Built on a 32.5B-parameter dense transformer, it achieves significantly enhanced performance on complex tasks—mathematics, coding, and multi-step reasoning—through extended chain-of-thought thinking. Available open-weight on Hugging Face, it delivers frontier reasoning in an efficient package.
Get Started
Model Specs
Released2025-03-05
Parameters32.5B
Context128k
ArchitectureDecoder Only