LLM Reference

Qwen3.8 Models by Alibaba

4 models2026Up to 1m ctxFrom $0.5/1M input

Last refreshed 2026-09-02. Next refresh: weekly.

Details

ResearcherAlibaba
Models4
Released2026
Max context1m

Capabilities

VisionAll models
MultimodalAll models
ReasoningAll models
JSON / Tool use3 of 4 models
Structured Outputs3 of 4 models

About

Qwen3.8 is Alibaba's August 2026 Qwen generation. It includes Qwen3.8-Max, a 2.4T-total / 95B-active MoE flagship on the public DashScope/QwenCloud API, and Qwen3.8-27B, a distinct Apache 2.0 dense vision-language model with official Hugging Face weights and a public Model Studio/QwenCloud API. Compare it for Coding, Agents, Long context, and Vision.

Decision facts

Best fit
agentcodingvision and multimodal work
Capability starting point
Qwen3.8-Max-0902 with 1m context and reasoning, JSON / Tool use, structured outputs, and multimodal inputs
Lowest tracked input
Qwen3.8-27B · $0.5/1M · Alibaba Cloud PAI-EAS
Closest related family
Tongyi DeepResearch

Current Variants

Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.

4 in view

Use when the workload needs 1m context, reasoning, and JSON / Tool use.

2026-091m contextreasoningJSON / Tool use

Use when the workload needs 262k context, reasoning, and multimodal inputs.

2026-08262k contextreasoningmultimodal inputs

Use when the workload needs 1m context, 27B parameters, and reasoning.

2026-081m context27B parametersreasoning

Use when the workload needs 1m context, reasoning, and JSON / Tool use.

2026-081m contextreasoningJSON / Tool use

Release Timeline

2 release groups
2026-09
1 current
Qwen3.8-Max-0902
1m contextreasoningJSON / Tool use
Current
2026-08
3 current
Qwen3.8-27B
1m context27B parametersreasoning
Current
Qwen3.8-Flash-Next
262k contextreasoningmultimodal inputs
Current
Qwen3.8-Max
1m contextreasoningJSON / Tool use
Current

Specifications(4 models)

Qwen3.8 model specifications comparison
ModelReleasedContextParametersVisionMultimodalReasoningJSON / Tool useStructured Outputs
Qwen3.8-Max-09022026-091m2.4TYesYesYesYesYes
Qwen3.8-Flash-Next2026-08262k125B total, 6B active (+51B n-gram embedding, 4B MTP)YesYesYesNoNo
Qwen3.8-27B2026-081m27BYesYesYesYesYes
Qwen3.8-Max2026-081m2.4T total, 95B activeYesYesYesYesYes

Available From(1 provider)

Pricing

Qwen3.8 model pricing by provider
ModelProviderInput / 1MOutput / 1MType
Qwen3.8-27BAlibaba Cloud PAI-EAS$0.5$3Serverless
Qwen3.8-MaxAlibaba Cloud PAI-EAS$2$6Serverless
Qwen3.8-Max-0902Alibaba Cloud PAI-EAS$2$6Serverless