Last refreshed 2026-06-01. Next refresh: weekly.
Why use Qwen3-32B on GroqCloud?
GroqCloud offers Qwen3-32B with pay-as-you-go pricing at $0.29/1M input tokens. Groq is a company specializing in AI inference technology, particularly with their flagship product, the Language Processing Unit (LPU™).
Compare Qwen3-32B across 7 providers to find the best fit for your use caseInput / 1M
$0.29
Output / 1M
$0.59
Cache
Not sourced
Batch
Not sourced
Setup recipe
Docs fallbackInstall
Use the provider REST API or SDKAuth
Create a provider API keyCall
model: qwen/qwen3-32bModel ID
qwen/qwen3-32bRequest example
Curated snippets for this provider are not sourced yet. Use GroqCloud documentation with model ID
qwen/qwen3-32b.Gotchas
- Use provider model ID "qwen/qwen3-32b", not the LLMReference slug "qwen3-32b".
Compare Qwen3-32B Across Providers
| Provider | Input (per 1M) | Output (per 1M) |
|---|---|---|
| Fireworks AI | $0.90 | $0.90 |
| GroqCloud | $0.29 | $0.59 |
| AWS Bedrock | $0.15 | $0.62 |
| OpenRouter | $0.08 | $0.24 |
| Vercel AI Gateway | $0.16 | $0.64 |
Pricing
| Type | Price (per 1M) |
|---|---|
| Input tokens | $0.29 |
| Output tokens | $0.59 |
Capabilities
Structured Outputs
About Qwen3-32B
Qwen3-32B is Alibaba's Qwen3 model. It offers a 40K-token context window.
Get Started
Model Specs
Released2025-04-29
Parameters32B
Context40k
ArchitectureDecoder Only