Ling-2.6-Flash
Ling-2.6-Flash is worth evaluating for rag, agents, and long context when its provider route and context window match the workload.
Use it for
- Teams evaluating rag, agents, and long context
- Workloads that can use a 262k context window
- Buyers comparing 2 tracked provider routes
Do not use it for
- Vision or document-understanding workloads
- Family
- Ling 2.6
- Released
- 2026-04-21
- Context
- 262k
- Parameters
- 104B (7.4B activated)
- Architecture
- Mixture of Experts
- Specialization
- general
- Openness
- Open source
- License
- Apache 2.0OSI-approvedCommercial use: permitted
- Weights
- Unknown
- Code
- Unknown
- Training
- Pretrained
InclusionAI is Ant Group's artificial general intelligence research lab, responsible for developing the Ling series of l
Cheapest of 2 routes · OpenRouter
About
InclusionAI's efficient 104B MoE instruct model with only 7.4B active parameters per token. Purpose-built for agentic workflows requiring fast responses and high token efficiency. Achieves 59.3% on GPQA Diamond. Nearly double the Artificial Analysis Intelligence Index score of comparable open-weight models. Available free on OpenRouter (inclusionai/ling-2.6-flash:free).
Ling-2.6-Flash is an open-source model in the Ling 2.6 family. The structured metadata tracks a 262k-token context window, function calling, tool use, and structured outputs. This page tracks provider routes through OpenRouter and Novita AI, with the cheapest tracked route listed at $0.01 input and $0.03 output per 1M tokens. No headline benchmark score is tracked for Ling-2.6-Flash yet.
Top use-case fit: coding, agents, and build tasks
RAG
Included by capability and metadata signals in the decision map.
Agents
Included by capability and metadata signals in the decision map.
Long context
Included by capability and metadata signals in the decision map.
Provider price ladder
Compare all 2Compare API pricing across 2 providers for input and output tokens, batch, and cached reads when available.
| Provider | Input / 1M | Output / 1M | Route |
|---|---|---|---|
| OpenRouter | $0.010 | $0.030 | Serverless |
| Novita AI | $0.100 | $0.300 | Serverless |
Capabilities
Benchmark peer barsfor RAG
No task-mapped benchmark peers are available for this model yet.
Migration checks
No linked migration route is available for this model yet.
Rankings & picks(1)
Compare Ling-2.6-Flash with other models
Comparison and alternatives
Browse all comparisons →Frequently asked questions
What is the context window of Ling-2.6-Flash?
Ling-2.6-Flash has a context window of 262k tokens.
How much does Ling-2.6-Flash cost?
Ling-2.6-Flash pricing ranges from $0.01/1M to $0.1/1M input tokens depending on the provider.
When was Ling-2.6-Flash released?
Ling-2.6-Flash was released on 2026-04-21.
Which providers offer Ling-2.6-Flash?
Ling-2.6-Flash is available from 2 providers: OpenRouter, Novita AI.
InclusionAI is Ant Group's artificial general intelligence research lab, responsible for developing the Ling series of l
Cheapest of 2 routes · OpenRouter