LLM Reference

Ling-2.6-Flash

Released
2026-04-21
Last refreshed
2026-06-29
Status
Researched 66d ago
Open sourceCommercial use: permittedRAGAgentsLong contextClassificationJSON / Tool use

Ling-2.6-Flash is worth evaluating for rag, agents, and long context when its provider route and context window match the workload.

Use it for

  • Teams evaluating rag, agents, and long context
  • Workloads that can use a 262k context window
  • Buyers comparing 2 tracked provider routes

Do not use it for

  • Vision or document-understanding workloads
Specifications
Family
Ling 2.6
Released
2026-04-21
Context
262k
Parameters
104B (7.4B activated)
Architecture
Mixture of Experts
Specialization
general
Openness
Open source
License
Apache 2.0OSI-approvedCommercial use: permitted
Weights
Unknown
Code
Unknown
Training
Pretrained
Created by

InclusionAI is Ant Group's artificial general intelligence research lab, responsible for developing the Ling series of l

Hangzhou, China
Founded 2023
Website
Pricing
Output / 1M
$0.030
Input / 1M
$0.010

Cheapest of 2 routes · OpenRouter

About

InclusionAI's efficient 104B MoE instruct model with only 7.4B active parameters per token. Purpose-built for agentic workflows requiring fast responses and high token efficiency. Achieves 59.3% on GPQA Diamond. Nearly double the Artificial Analysis Intelligence Index score of comparable open-weight models. Available free on OpenRouter (inclusionai/ling-2.6-flash:free).

Ling-2.6-Flash is an open-source model in the Ling 2.6 family. The structured metadata tracks a 262k-token context window, function calling, tool use, and structured outputs. This page tracks provider routes through OpenRouter and Novita AI, with the cheapest tracked route listed at $0.01 input and $0.03 output per 1M tokens. No headline benchmark score is tracked for Ling-2.6-Flash yet.

Top use-case fit: coding, agents, and build tasks

RAG

Included by capability and metadata signals in the decision map.

Agents

Included by capability and metadata signals in the decision map.

Long context

Included by capability and metadata signals in the decision map.

Provider price ladder

Compare all 2

Compare API pricing across 2 providers for input and output tokens, batch, and cached reads when available.

ProviderInput / 1MOutput / 1MRoute
OpenRouter$0.010$0.030
Serverless
Novita AI$0.100$0.300
Serverless

Capabilities

Function CallingTool UseStructured Outputs

Benchmark peer barsfor RAG

No task-mapped benchmark peers are available for this model yet.

Migration checks

No linked migration route is available for this model yet.

Compare Ling-2.6-Flash with other models

Frequently asked questions

What is the context window of Ling-2.6-Flash?

Ling-2.6-Flash has a context window of 262k tokens.

How much does Ling-2.6-Flash cost?

Ling-2.6-Flash pricing ranges from $0.01/1M to $0.1/1M input tokens depending on the provider.

When was Ling-2.6-Flash released?

Ling-2.6-Flash was released on 2026-04-21.

Which providers offer Ling-2.6-Flash?

Ling-2.6-Flash is available from 2 providers: OpenRouter, Novita AI.