LLM Reference

Qwen3 VL 30B A3B Instruct

Released
2025-09-18
Last refreshed
2026-06-29
Status
Researched 71d ago
Open sourceCommercial use: permittedMultimodalRAGAgentsLong contextVisionJSON / Tool use

Qwen3 VL 30B A3B Instruct is worth evaluating for rag, agents, and long context when its provider route and context window match the workload.

Use it for

  • Teams evaluating rag, agents, and long context
  • Workloads that can use a 128k context window
  • Buyers comparing 1 tracked provider route

Do not use it for

  • Workloads where another current model has stronger sourced task evidence
Specifications
Family
Qwen3-VL
Released
2025-09-18
Context
128k
Parameters
30B
Architecture
Mixture of Experts
Specialization
general
Openness
Open source
License
Apache 2.0OSI-approvedCommercial use: permitted
Weights
Available
Code
Unknown
Training
Pretrained
Created by

AI research institute of Alibaba Group.

Hangzhou, Zhejiang, China
Founded 2017
Website
Pricing
Output / 1M
$0.700
Input / 1M
$0.200

Cheapest of 1 route · Novita AI

About

Qwen3-VL-30B-A3B-Instruct is a multimodal MoE model from Alibaba unifying text generation with visual understanding for images, charts, and documents at 128K context.

Qwen3 VL 30B A3B Instruct is an open-source model in the Qwen3-VL family. The structured metadata tracks a 128k-token context window, multimodal input, function calling, tool use, and structured outputs. This page tracks provider routes through Novita AI, with the cheapest tracked route listed at $0.2 input and $0.7 output per 1M tokens. Headline tracked benchmarks include MMMU Pro 60.4.

Top use-case fit: coding, agents, and build tasks

RAG

Included by capability and metadata signals in the decision map.

Agents

Included by capability and metadata signals in the decision map.

Long context

Included by capability and metadata signals in the decision map.

Provider price ladder

Compare API pricing across 1 providers for input and output tokens, batch, and cached reads when available.

ProviderInput / 1MOutput / 1MRoute
Novita AI$0.200$0.700
Serverless

Capabilities

VisionMultimodalFunction CallingTool UseStructured Outputs

Benchmark peer barsfor RAG

No task-mapped benchmark peers are available for this model yet.

Benchmark scores(1)

Scores are benchmark-specific and are direction-aware: the same numeric gap can mean very different outcomes across suites. Use the leaderboard context and this model's provider route to decide whether the winning margin is meaningful for your workload.
BenchmarkScoreVersionEvaluationSource
MMMU Pro60.4LLM-Stats aggregator, instruct-onlyObserved 2026-06-07Source

Migration checks

No linked migration route is available for this model yet.

Frequently asked questions

What is the context window of Qwen3 VL 30B A3B Instruct?

Qwen3 VL 30B A3B Instruct has a context window of 128k tokens.

How much does Qwen3 VL 30B A3B Instruct cost?

Qwen3 VL 30B A3B Instruct is available at $0.2/1M input tokens through Novita AI.

When was Qwen3 VL 30B A3B Instruct released?

Qwen3 VL 30B A3B Instruct was released on 2025-09-18.

Which providers offer Qwen3 VL 30B A3B Instruct?

Qwen3 VL 30B A3B Instruct is available from 1 provider: Novita AI.

What benchmarks has Qwen3 VL 30B A3B Instruct been tested on?

Qwen3 VL 30B A3B Instruct has been evaluated on 1 benchmark, including MMMU Pro.