Qwen3.6-35B-A3B on Novita AI

Qwen3.6 · Alibaba

ServerlessOpen Source

Last refreshed 2026-06-29. Next refresh: weekly.

Why use Qwen3.6-35B-A3B on Novita AI?

Novita AI offers Qwen3.6-35B-A3B with pay-as-you-go pricing at $0.25/1M input tokens. Novita AI offers a GPU-based inference API for image, video, and language model generation with a broad catalog of open-source models.

Compare Qwen3.6-35B-A3B across 2 providers to find the best fit for your use case
Input / 1M
$0.248
Output / 1M
$1.49
Cache
Not sourced
Batch
Not sourced

Setup recipe

Docs fallback
Install
Use the provider REST API or SDK
Auth
Create a provider API key
Call
model: qwen3.6-35b-a3b
Model ID
qwen3.6-35b-a3b

Request example

Curated snippets for this provider have not been sourced yet.

Gotchas

No curated gotchas have been sourced for this exact provider/model route yet.

Compare Qwen3.6-35B-A3B Across Providers

ProviderInput (per 1M)Output (per 1M)
OpenRouter$0.15$1.00
Novita AI$0.25$1.49

Pricing

TypePrice (per 1M)
Input tokens$0.25
Output tokens$1.49

Capabilities

VisionMultimodalJSON / Tool use

About Qwen3.6-35B-A3B

Qwen3.6-35B-A3B is an open-weight multimodal MoE model with 35B total parameters and 3B activated per token, released April 2026. It features a hybrid architecture combining Gated DeltaNet linear attention and standard Gated Attention with 256 total experts (8 routed + 1 shared), and includes a vision encoder for image and video understanding. Optimized for agentic coding, long-context reasoning, and visual tasks; supports 256K native context (extensible to ~1M via YaRN) with integrated thinking mode for multi-turn agent interactions.

Get Started

Model Specs

Released2026-04-16
Parameters35B
Context262k
ArchitectureMixture of Experts

Related Models on Novita AI