Qwen3-32B on Novita AI

Qwen3 · Alibaba

ServerlessOpen Source

Last refreshed 2026-06-29. Next refresh: weekly.

Why use Qwen3-32B on Novita AI?

Novita AI offers Qwen3-32B with pay-as-you-go pricing at $0.10/1M input tokens. Novita AI offers a GPU-based inference API for image, video, and language model generation with a broad catalog of open-source models.

Compare Qwen3-32B across 7 providers to find the best fit for your use case
Input / 1M
$0.10
Output / 1M
$0.45
Cache
Not sourced
Batch
Not sourced

Setup recipe

Docs fallback
Install
Use the provider REST API or SDK
Auth
Create a provider API key
Call
model: qwen3-32b-fp8
Model ID
qwen3-32b-fp8

Request example

Curated snippets for this provider have not been sourced yet.

Gotchas

  • Use provider model ID "qwen3-32b-fp8", not the LLMReference slug "qwen3-32b".

Compare Qwen3-32B Across Providers

ProviderInput (per 1M)Output (per 1M)
Fireworks AI$0.90$0.90
GroqCloud$0.29$0.59
AWS Bedrock$0.15$0.62
OpenRouter$0.08$0.24
Vercel AI Gateway$0.16$0.64
View all 7 providers →

Pricing

TypePrice (per 1M)
Input tokens$0.10
Output tokens$0.45

Capabilities

Structured Outputs

About Qwen3-32B

Qwen3-32B is Alibaba's Qwen3 model. It offers a 40K-token context window.

Get Started