LLM Reference

Qwen3.8-Omni-Flash

Released
2026-09-18
Last refreshed
2026-09-17
Status
Researched 1d ago
ProprietaryCommercial use: conditionalMultimodalRAGAgentsLong contextVisionJSON / Tool useMultimodalAgents

Qwen3.8-Omni-Flash is worth evaluating for rag, agents, and long context when its provider route and context window match the workload.

Use it for

  • Teams evaluating rag, agents, and long context
  • Workloads that can use a 1m context window
  • Buyers comparing 1 tracked provider route

Do not use it for

  • Workloads where another current model has stronger sourced task evidence
Specifications
Family
Qwen3.8
Released
2026-09-18
Context
1m
Max output
131,072
Specialization
general
Openness
Proprietary
License
ProprietaryCommercial use: conditional
Weights
Not released
Code
Unknown
Training
Pretrained
Created by

AI research institute of Alibaba Group.

Hangzhou, Zhejiang, China
Founded 2017
Website
Pricing
Output / 1M
$0.470
Input / 1M
$0.150

Cheapest of 1 route · Alibaba Cloud PAI-EAS · cache read $0.016

About

Qwen3.8-Omni-Flash is Alibaba Qwen's first omni-modal model built around agentic capabilities, announced 2026-09-18 (MCP-verified @Alibaba_Qwen tip 2100785962414702599 + first-party QwenCloud / Model Studio Qwen-Omni docs). Native text, image, audio, and video input with text-only output (do not set audio modalities). 1M-token context; QwenCloud lists max input 991K / max output 131K (thinking: max input 983K / max output 131K / max reasoning 262K). Tip + QwenCloud: audio-video understanding with reasoning and tool use for agentic workflows (vlog auto-edit, short-video translate, movie recaps); two-/four-channel spatial audio; DashScope + OpenAI-compatible protocols; companion open-source Qwen-MM-Plugins (Qwen-Live Harness coming soon).

Top use-case fit: coding, agents, and build tasks

RAG

Included by capability and metadata signals in the decision map.

Agents

Included by capability and metadata signals in the decision map.

Long context

Included by capability and metadata signals in the decision map.

Provider price ladder

Compare API pricing across 1 providers for input and output tokens, batch, and cached reads when available.

ProviderInput / 1MOutput / 1MCacheRoute
Alibaba Cloud PAI-EAS$0.150$0.470read $0.016
Serverless

Capabilities

VisionMultimodalReasoningJSON / Tool useStructured OutputsPrompt CachingAudio

Benchmark peer barsfor RAG

No task-mapped benchmark peers are available for this model yet.

Migration checks

No linked migration route is available for this model yet.

API versions

qwen3.8-omni-flashQwen3.8-Omni-Flash