LLM Reference
OpenRouter

Using Ling 3.0 Flash on OpenRouter

Implementation guide · Ling 3.0 · InclusionAI

ServerlessOpen Source

OpenRouter exposes Ling 3.0 Flash through model ID inclusionai/ling-3.0-flash. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-09-07. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at OpenRouter and generate an API key.
  2. 2
    Use the OpenRouter SDK or REST API to call inclusionai/ling-3.0-flash — see the documentation for request format.
  3. 3
    You'll be billed $0.02/1M input, $0.06/1M output tokens. See full pricing.

Code Examples

See OpenRouter documentation for integration details.

Pricing on OpenRouter

TypePrice (per 1M)
Input tokens$0.02
Output tokens$0.06

Capabilities

ReasoningJSON / Tool useStructured Outputs

About Ling 3.0 Flash

Ling-3.0-flash is InclusionAI's next-generation hybrid-linear MoE instruct model with 124B total and 5.1B active parameters per token, 262K context, and agentic training across coding, research, and tool-use environments. MIT-licensed weights on Hugging Face (inclusionAI/Ling-3.0-flash); available on OpenRouter as inclusionai/ling-3.0-flash.

Model Specs

Released2026-08-02
Parameters124B (5.1B activated)
Context262k
ArchitectureMixture of Experts

Provider

OpenRouter
OpenRouter

OpenRouter, Inc.

New York, NY, USA