LLM Reference

Qwen3 Models by Alibaba

AlibabaApache 2.0Open source
19 models2025Up to 1m ctxFrom $0.035/1M input

Last refreshed 2026-06-01. Next refresh: weekly.

Details

ResearcherAlibaba
LicenseApache 2.0OSI-approved
Commercial useCommercial use: permitted
Models19
Released2025
Max context1m

Capabilities

Vision1 of 19 models
Multimodal1 of 19 models
Reasoning2 of 19 models
JSON / Tool use2 of 19 models
Structured Outputs8 of 19 models

About

Qwen3 is a family of 19 AI models by Alibaba, released in 2025.

Decision facts

Best fit
vision and multimodal workreasoningJSON / Tool use
Capability starting point
Qwen3-Max with 262k context and JSON / Tool use, structured outputs, and multimodal inputs
Lowest tracked input
Qwen3-8B · $0.035/1M · Novita AI
Closest related family
Tongyi DeepResearch

Current Variants

Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.

19 in view
Qwen3-105BCurrent

Use when the workload needs 128k context, 105B parameters, and JSON / Tool use.

2025-12128k context105B parametersJSON / Tool use

Use when the workload needs structured outputs.

2025-12structured outputs
Qwen-PlusCurrent

Use when the workload needs 1m context.

2025-111m context
Qwen3-8BCurrent

Use when the workload needs 128k context, 8B parameters, and structured outputs.

2025-08128k context8B parametersstructured outputs

Use when the workload needs 128k context and 8B parameters.

2025-08128k context8B parameters
Qwen3-70BCurrent

Use when the workload needs 128k context and 70B parameters.

2025-08128k context70B parameters

Use when the workload needs 128k context and 70B parameters.

2025-08128k context70B parameters
Qwen-FlashCurrent

Use when the workload needs 1m context.

2025-081m context

Use when the workload needs 128k context, 235B parameters, and structured outputs.

2025-04128k context235B parametersstructured outputs
Qwen3-32BCurrent

Use when the workload needs 40k context, 32B parameters, and structured outputs.

2025-0440k context32B parametersstructured outputs
Qwen3-MaxCurrent

Use when the workload needs 262k context, JSON / Tool use, and structured outputs.

2025-04262k contextJSON / Tool usestructured outputs

Use when the workload needs 128k context, 30B parameters, and structured outputs.

2025-04128k context30B parametersstructured outputs
Qwen3-9BCurrent

Use when the workload needs 256k context, 9B parameters, and structured outputs.

2025-04256k context9B parametersstructured outputs

Use when the workload needs 160k context, 8B parameters, and reasoning.

2025-01160k context8B parametersreasoning
Qwen3-0.6BCurrent

Use when the workload needs 40k context and 600M parameters.

2025-0140k context600M parameters
Qwen3-14BCurrent

Use when the workload needs 40k context, 14B parameters, and structured outputs.

2025-0140k context14B parametersstructured outputs
Qwen3-1.7BCurrent

Use when the workload needs 40k context and 1.7B parameters.

2025-0140k context1.7B parameters
Qwen3-4BCurrent

Use when the workload needs 40k context and 4B parameters.

2025-0140k context4B parameters

Use when the workload needs 160k context, 671B parameters, and reasoning.

2025-01160k context671B parametersreasoning

Release Timeline

5 release groups
2025-12
2 current
Qwen3-105B
128k context105B parametersJSON / Tool use
Current
Qwen3-Next-80B-A3B
structured outputs
Current
2025-11
1 current
Qwen-Plus
1m context
Current
2025-08
5 current
Qwen-Flash
1m context
Current
Qwen3-70B
128k context70B parameters
Current
Qwen3-70B-Instruct
128k context70B parameters
Current
Qwen3-8B
128k context8B parametersstructured outputs
Current
Qwen3-8B-Instruct
128k context8B parameters
Current
2025-04
5 current
Qwen3-235B-A22B
128k context235B parametersstructured outputs
Current
Qwen3-30B-A3B
128k context30B parametersstructured outputs
Current
Qwen3-32B
40k context32B parametersstructured outputs
Current
Qwen3-9B
256k context9B parametersstructured outputs
Current
Qwen3-Max
262k contextJSON / Tool usestructured outputs
Current
2025-01
6 current
DeepSeek R1 0528 Distill Qwen3-8B
160k context8B parametersreasoning
Current
DeepSeek R1 0528 Qwen3-8B
160k context671B parametersreasoning
Current
Qwen3-0.6B
40k context600M parameters
Current
Qwen3-1.7B
40k context1.7B parameters
Current
Qwen3-14B
40k context14B parametersstructured outputs
Current
Qwen3-4B
40k context4B parameters
Current

Specifications(19 models)

Qwen3 model specifications comparison
ModelReleasedContextParametersVisionMultimodalReasoningJSON / Tool useStructured Outputs
Qwen3-105B2025-12128k105BNoNoNoYesNo
Qwen3-Next-80B-A3B2025-1280B (3B active)NoNoNoNoYes
Qwen-Plus2025-111mNoNoNoNoNo
Qwen3-8B2025-08128k8BNoNoNoNoYes
Qwen3-8B-Instruct2025-08128k8BNoNoNoNoNo
Qwen3-70B2025-08128k70BNoNoNoNoNo
Qwen3-70B-Instruct2025-08128k70BNoNoNoNoNo
Qwen-Flash2025-081mNoNoNoNoNo
Qwen3-235B-A22B2025-04128k235BNoNoNoNoYes
Qwen3-32B2025-0440k32BNoNoNoNoYes
Qwen3-Max2025-04262kYesYesNoYesYes
Qwen3-30B-A3B2025-04128k30BNoNoNoNoYes
Qwen3-9B2025-04256k9BNoNoNoNoYes
DeepSeek R1 0528 Distill Qwen3-8B2025-01160k8BNoNoYesNoNo
Qwen3-0.6B2025-0140k0.6BNoNoNoNoNo
Qwen3-14B2025-0140k14BNoNoNoNoYes
Qwen3-1.7B2025-0140k1.7BNoNoNoNoNo
Qwen3-4B2025-0140k4BNoNoNoNoNo
DeepSeek R1 0528 Qwen3-8B2025-01160k671BNoNoYesNoNo

Pricing

Qwen3 model pricing by provider
ModelProviderInput / 1MOutput / 1MType
Qwen3-8BNovita AI$0.035$0.138Serverless
Qwen3-9BDeepInfra$0.04$0.2Serverless
Qwen3-8BOpenRouter$0.05$0.4Serverless
Qwen3-30B-A3BCloudflare Workers AI$0.051$0.335Serverless
Qwen3-14BOpenRouter$0.06$0.24Serverless
DeepSeek R1 0528 Qwen3-8BNovita AI$0.06$0.09Serverless
Qwen3-30B-A3BOpenRouter$0.08$0.28Serverless
Qwen3-32BOpenRouter$0.08$0.24Serverless
Qwen3-30B-A3BVercel AI Gateway$0.08$0.29Serverless
Qwen3-235B-A22BNovita AI$0.09$0.58Serverless
Qwen3-30B-A3BNovita AI$0.09$0.45Serverless
Qwen3-Next-80B-A3BOpenRouter$0.0975$0.78Serverless
Qwen3-0.6BFireworks AI$0.1$0.1Serverless
Qwen3-1.7BFireworks AI$0.1$0.1Serverless
Qwen3-30B-A3BAWS Bedrock$0.1$0.3Serverless
Qwen3-32BNovita AI$0.1$0.45Serverless
Qwen3-14BNextBit$0.1$0.24Serverless
Qwen3-14BVercel AI Gateway$0.12$0.24Serverless
Qwen3-30B-A3BNextBit$0.14$0.55Serverless
Qwen3-32BAWS Bedrock$0.15$0.62Serverless
Qwen3-Next-80B-A3BAWS Bedrock$0.15$1.2Serverless
Qwen3-Next-80B-A3BNovita AI$0.15$1.5Serverless
Qwen3-32BVercel AI Gateway$0.16$0.64Serverless
DeepSeek R1 0528 Distill Qwen3-8BFireworks AI$0.2$0.2Serverless
Qwen3-14BFireworks AI$0.2$0.2Serverless
Qwen3-4BFireworks AI$0.2$0.2Serverless
Qwen3-8BFireworks AI$0.2$0.2Serverless
DeepSeek R1 0528 Qwen3-8BFireworks AI$0.2$0.2Serverless
Qwen-FlashAlibaba Cloud PAI-EAS$0.25$2Serverless
Qwen3-32BGroqCloud$0.29$0.59Serverless
Qwen3-235B-A22BAWS Bedrock$0.4$1.2Serverless
Qwen3-235B-A22BOpenRouter$0.455$1.82Serverless
Qwen3-30B-A3BFireworks AI$0.5$0.5Serverless
Qwen3-MaxOpenRouter$0.78$3.9Serverless
Qwen3-32BFireworks AI$0.9$0.9Serverless
Qwen3-235B-A22BFireworks AI$1.2$1.2Serverless
Qwen-PlusAlibaba Cloud PAI-EAS$1.2$3.6Serverless
Qwen3-MaxVercel AI Gateway$1.2$6Serverless
Qwen3-MaxNovita AI$2.11$8.45Serverless

Popular comparisons in this family

Models(19)