Last refreshed 2026-09-07. Next refresh: weekly.
Why use Ling 3.0 Flash on OpenRouter?
OpenRouter offers Ling 3.0 Flash with pay-as-you-go pricing at $0.02/1M input tokens. OpenRouter is a multi-provider LLM aggregator offering unified API access to 300+ models from all major labs and emerging providers, with automatic failover for reliability.
Input / 1M
$0.021
Output / 1M
$0.063
Cache
read $0.0042
Batch
Not sourced
Setup recipe
Docs fallbackInstall
Use the provider REST API or SDKAuth
Create a provider API keyCall
model: inclusionai/ling-3.0-flashModel ID
inclusionai/ling-3.0-flashRequest example
Curated snippets for this provider are not sourced yet. Use OpenRouter documentation with model ID
inclusionai/ling-3.0-flash.Gotchas
- Use provider model ID "inclusionai/ling-3.0-flash", not the LLMReference slug "ling-3.0-flash".
Pricing
| Type | Price (per 1M) |
|---|---|
| Input tokens | $0.02 |
| Output tokens | $0.06 |
Capabilities
ReasoningJSON / Tool useStructured Outputs
About Ling 3.0 Flash
Ling-3.0-flash is InclusionAI's next-generation hybrid-linear MoE instruct model with 124B total and 5.1B active parameters per token, 262K context, and agentic training across coding, research, and tool-use environments. MIT-licensed weights on Hugging Face (inclusionAI/Ling-3.0-flash); available on OpenRouter as inclusionai/ling-3.0-flash.
Get Started
Model Specs
Released2026-08-02
Parameters124B (5.1B activated)
Context262k
ArchitectureMixture of Experts