LLM Reference
Alibaba Cloud PAI-EAS

Using Qwen-Flash on Alibaba Cloud PAI-EAS

Implementation guide · Qwen3 · Alibaba

ServerlessOpen Source

Alibaba Cloud PAI-EAS exposes Qwen-Flash through model ID qwen-flash. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-04-21. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at Alibaba Cloud PAI-EAS and generate an API key.
  2. 2
    Use the Alibaba Cloud PAI-EAS SDK or REST API to call qwen-flash — see the documentation for request format.
  3. 3
    You'll be billed $0.25/1M input, $2.00/1M output tokens. See full pricing.

Code Examples

See Alibaba Cloud PAI-EAS documentation for integration details.

Pricing on Alibaba Cloud PAI-EAS

TypePrice (per 1M)
Input tokens$0.25
Output tokens$2.00

Capabilities

No model capability flags are currently sourced.

About Qwen-Flash

Qwen-Flash is a Qwen3 series Flash model that seamlessly integrates thinking and non-thinking modes switchable mid-dialogue, excelling at complex thinking tasks with significant improvements in instruction adherence and text understanding. It supports 1M context length with tiered pricing based on context length.

Model Specs

Released2025-08-01
Context1m

More Models on Alibaba Cloud PAI-EAS

Provider

Alibaba Cloud PAI-EAS
Alibaba Cloud PAI-EAS

Alibaba Cloud

Hangzhou, Zhejiang, China