Last refreshed 2026-10-02. Next refresh: weekly.
Why use Ling-3.1-flash on OpenRouter?
OpenRouter offers Ling-3.1-flash with free input token pricing. OpenRouter is a multi-provider LLM aggregator offering unified API access to 300+ models from all major labs and emerging providers, with automatic failover for reliability.
Compare Ling-3.1-flash across 2 providers to find the best fit for your use caseSetup recipe
Docs fallbackUse the provider REST API or SDKCreate a provider API keymodel: inclusionai/ling-3.1-flashinclusionai/ling-3.1-flashRequest example
inclusionai/ling-3.1-flash.Gotchas
- Use provider model ID "inclusionai/ling-3.1-flash", not the LLMReference slug "ling-3.1-flash".
Compare Ling-3.1-flash Across Providers
| Provider | Input (per 1M) | Output (per 1M) |
|---|---|---|
| OpenRouter | Free | Free |
| Vercel AI Gateway | Free | Free |
Pricing
| Type | Price (per 1M) |
|---|---|
| Input tokens | Free |
| Output tokens | Free |
Capabilities
About Ling-3.1-flash
Ling-3.1-flash is InclusionAI's (Ant Group 蚂蚁百灵 / Ant Ling) hybrid reasoning mixture-of-experts Flash model announced 2026-09-30. Authoritative Chinese IT Home coverage: ~560B total parameters with ~25B activated per token; designed context window up to 1M tokens; two-week free trial capped at 256K context, with paid 1M context and planned open-source weights after the trial. TechNode English summary of the same launch positions it for agent tasks, search, office software, and specialist applications.