LLM Reference

Qwen3.8-Max

Released
2026-08-03
Last refreshed
2026-08-24
Status
Researched 1d ago
Open weightsCommercial use: conditionalMultimodalCodingRAGAgentsLong contextVisionJSON / Tool use

Qwen3.8-Max is worth evaluating for coding, rag, and agents when its provider route and context window match the workload.

Use it for

  • Teams evaluating coding, rag, and agents
  • Workloads that can use a 1m context window
  • Buyers comparing 1 tracked provider route

Do not use it for

  • Workloads where another current model has stronger sourced task evidence
Specifications
Family
Qwen3.8
Released
2026-08-03
Context
1m
Max output
131,072
Parameters
2.4T total, 95B active
Architecture
Mixture of Experts
Openness
Open weights
License
Qwen3.8-Max LicenseCommercial use: conditional
Weights
Available
Code
Unknown
Created by

AI research institute of Alibaba Group.

Hangzhou, Zhejiang, China
Founded 2017
Website
Pricing
Output / 1M
$6.00
Input / 1M
$2.00

Cheapest of 1 route · Alibaba Cloud PAI-EAS · cache read $0.170

About

Qwen3.8-Max is a 2.4-trillion-parameter MoE flagship delivering a comprehensive leap in coding and professional work. It autonomously codes and delivers complete projects spanning 10+ days, handles hundreds of specialized tasks across legal, financial, design, and other professional domains, and produces production-grade results end-to-end in a single conversation. Native visual understanding runs through the full cycle of planning, execution, and verification, enabling deep semantic analysis of ultra-long documents and extended video content.

Top use-case fit: coding, agents, and build tasks

Coding

Q/$ D

1 relevant benchmark in the decision map.

RAG

Included by capability and metadata signals in the decision map.

Agents

Included by capability and metadata signals in the decision map.

Provider price ladder

Compare API pricing across 1 providers for input and output tokens, batch, and cached reads when available.

ProviderInput / 1MOutput / 1MCacheRoute
Alibaba Cloud PAI-EAS$2.00$6.00read $0.170
Serverless

Capabilities

VisionMultimodalReasoningJSON / Tool useStructured OutputsPrompt Caching

Benchmark peer barsfor Coding

Benchmark scores(5)

Scores are benchmark-specific and are direction-aware: the same numeric gap can mean very different outcomes across suites. Use the leaderboard context and this model's provider route to decide whether the winning margin is meaningful for your workload.
BenchmarkScoreVersionEvaluationSource
Terminal-Bench 2.186.6Terminal Bench 2.1Observed 2026-08-12Source
SWE-bench Pro67.7SWE-bench ProObserved 2026-08-12Source
DeepSWE 1.156.6DeepSWE 1.1Observed 2026-08-12Source
AutomationBench27.3Automation-Bench (Pass@1)Observed 2026-08-12Source
Google-Proof Q&A92.6GPQA DiamondObserved 2026-08-12Source

Migration checks

No linked migration route is available for this model yet.

API versions

qwen3.8-max