Using Ling-2.6-Flash on OpenRouter
Implementation guide · Ling 2.6 · InclusionAI
ServerlessOpen Source
OpenRouter exposes Ling-2.6-Flash through model ID inclusionai/ling-2.6-flash. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.
Last refreshed 2026-06-15. Next refresh: weekly.
Quick Start
- 1
- 2Use the OpenRouter SDK or REST API to call
inclusionai/ling-2.6-flash— see the documentation for request format. - 3
Code Examples
Pricing on OpenRouter
| Type | Price (per 1M) |
|---|---|
| Input tokens | $0.01 |
| Output tokens | $0.03 |
Capabilities
JSON / Tool useStructured Outputs
About Ling-2.6-Flash
InclusionAI's efficient 104B MoE instruct model with only 7.4B active parameters per token. Purpose-built for agentic workflows requiring fast responses and high token efficiency. Achieves 59.3% on GPQA Diamond. Nearly double the Artificial Analysis Intelligence Index score of comparable open-weight models. Available free on OpenRouter (inclusionai/ling-2.6-flash:free).
Model Specs
Released2026-04-21
Parameters104B (7.4B activated)
Context262k
ArchitectureMixture of Experts