Qwen3-Coder-480B-A35B-Instruct
- Family
- Qwen3-Coder
- Released
- 2025-07-22
- Context
- 262k
- Parameters
- 480B total, 35B active
- Architecture
- Mixture of Experts
- Specialization
- code
- Openness
- Open source
- License
- Apache 2.0OSI-approvedCommercial use: permitted
- Weights
- Available
- Code
- Unknown
- Training
- Pretrained
Cheapest of 7 routes · OpenRouter · cache read $0.100
About
Qwen3-Coder-480B-A35B-Instruct is Alibaba's flagship open-source code generation and agentic model, released July 22, 2025 under the Apache 2.0 license. The model has 480 billion total parameters with 35 billion active parameters per token, organized across 62 transformer layers with 160 specialized expert networks and 8 experts activated per token. It uses Grouped Query Attention (GQA) with 96 query heads and 8 key-value heads and supports a native context window of 262,144 tokens, extendable to 1 million tokens via YaRN position scaling.
Top use-case fit: coding, agents, and build tasks
Coding
Q/$ CAgents
Q/$ BProvider price ladder
Compare all 7Compare API pricing across 4 providers for input and output tokens, batch, and cached reads when available.
| Provider | Input / 1M | Output / 1M | Cache | Route |
|---|---|---|---|---|
| OpenRouter | $0.300 | $1.00 | read $0.100 | Serverless |
| Novita AI | $0.380 | $1.55 | - | Serverless |
| GCP Vertex AI | $0.220 | $1.80 | - | Serverless |
| Vercel AI Gateway | $1.50 | $7.50 | read $0.300 | Serverless |
Available via routers & gateways(15)
LiteLLM
GatewayOpen-source Python SDK and proxy server that unifies 100+ LLM APIs behind a single OpenAI-compatible interface, with load balancing, cost tracking, and configurable failover.
OpenRouter
HybridUnified hybrid gateway to 400+ models from 60+ providers via a single OpenAI-compatible API, with optional auto-routing that selects the best model per prompt.
Portkey
GatewayProduction AI gateway routing to 1,600+ LLMs with failover, load balancing, semantic caching, and guardrails; Apache 2.0 core is fully self-hostable with the complete feature set.
AIRouter
RouterCommercial LLM router that analyzes incoming requests and routes to the optimal model for cost/quality/latency via a drop-in OpenAI-compatible API, with a privacy-preserving embedding mode that avoids sending prompt content.
Amazon Bedrock Intelligent Prompt Routing
RouterAWS Bedrock's native intelligent prompt router that routes prompts between Anthropic Claude model tiers (Haiku/Sonnet) based on predicted task complexity, with no extra per-routing charge.
Helicone
GatewayObservability-first AI gateway with routing, caching, rate limiting, and request tracing; Apache 2.0 open-source core with a managed hosted tier for logging and analytics.
Capabilities
Benchmark peer barsfor Coding
Benchmark scores(3)
| Benchmark | Score | Version | Evaluation | Source |
|---|---|---|---|---|
| Berkeley Function Calling Leaderboard v3 | 68.7 | Berkeley Function Calling Leaderboard (BFCL v3)Observed 2026-04-12 | — | Source |
| SWE-bench Pro | 38.7 | Scale AI standardized SWE-bench ProObserved 2026-06-23 | — | Source |
| SWE-bench Verified | 66.5 | Nebius/OpenHands independent SWE-bench VerifiedObserved 2026-06-23 | — | Source |
Migration checks
No linked migration route is available for this model yet.
Rankings & picks(1)
Cheapest of 7 routes · OpenRouter · cache read $0.100