Using Ling 3.0 Flash on OpenRouter
Implementation guide · Ling 3.0 · InclusionAI
ServerlessOpen Source
OpenRouter exposes Ling 3.0 Flash through model ID inclusionai/ling-3.0-flash. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.
Last refreshed 2026-09-07. Next refresh: weekly.
Quick Start
- 1
- 2Use the OpenRouter SDK or REST API to call
inclusionai/ling-3.0-flash— see the documentation for request format. - 3
Code Examples
Pricing on OpenRouter
| Type | Price (per 1M) |
|---|---|
| Input tokens | $0.02 |
| Output tokens | $0.06 |
Capabilities
ReasoningJSON / Tool useStructured Outputs
About Ling 3.0 Flash
Ling-3.0-flash is InclusionAI's next-generation hybrid-linear MoE instruct model with 124B total and 5.1B active parameters per token, 262K context, and agentic training across coding, research, and tool-use environments. MIT-licensed weights on Hugging Face (inclusionAI/Ling-3.0-flash); available on OpenRouter as inclusionai/ling-3.0-flash.
Model Specs
Released2026-08-02
Parameters124B (5.1B activated)
Context262k
ArchitectureMixture of Experts