Last refreshed 2026-06-29. Next refresh: weekly.
Why use Qwen3.6-35B-A3B on Novita AI?
Novita AI offers Qwen3.6-35B-A3B with pay-as-you-go pricing at $0.25/1M input tokens. Novita AI offers a GPU-based inference API for image, video, and language model generation with a broad catalog of open-source models.
Compare Qwen3.6-35B-A3B across 2 providers to find the best fit for your use caseSetup recipe
Docs fallbackUse the provider REST API or SDKCreate a provider API keymodel: qwen3.6-35b-a3bqwen3.6-35b-a3bRequest example
Gotchas
No curated gotchas have been sourced for this exact provider/model route yet.
Compare Qwen3.6-35B-A3B Across Providers
| Provider | Input (per 1M) | Output (per 1M) |
|---|---|---|
| OpenRouter | $0.15 | $1.00 |
| Novita AI | $0.25 | $1.49 |
Pricing
| Type | Price (per 1M) |
|---|---|
| Input tokens | $0.25 |
| Output tokens | $1.49 |
Capabilities
About Qwen3.6-35B-A3B
Qwen3.6-35B-A3B is an open-weight multimodal MoE model with 35B total parameters and 3B activated per token, released April 2026. It features a hybrid architecture combining Gated DeltaNet linear attention and standard Gated Attention with 256 total experts (8 routed + 1 shared), and includes a vision encoder for image and video understanding. Optimized for agentic coding, long-context reasoning, and visual tasks; supports 256K native context (extensible to ~1M via YaRN) with integrated thinking mode for multi-turn agent interactions.