Using Ling-3.0-flash-VL on OpenRouter
Implementation guide · Ling 3.0 · InclusionAI
ServerlessOpen Source
OpenRouter exposes Ling-3.0-flash-VL through model ID inclusionai/ling-3.0-flash-vl. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.
Last refreshed 2026-09-08. Next refresh: weekly.
Quick Start
- 1
- 2Use the OpenRouter SDK or REST API to call
inclusionai/ling-3.0-flash-vl— see the documentation for request format. - 3
Code Examples
Pricing on OpenRouter
| Type | Price (per 1M) |
|---|---|
| Input tokens | $0.06 |
| Output tokens | $0.18 |
Capabilities
VisionMultimodalReasoningJSON / Tool useStructured Outputs
About Ling-3.0-flash-VL
Ling-3.0-flash-VL is InclusionAI's vision-language fine-tune of Ling-3.0-flash (124B total / 5.5B activated MoE) for image and video understanding, reasoning, function calling, and agentic tool use. Native config context is 131,072 tokens; the Hugging Face model card documents YaRN extension up to 1M tokens. MIT-licensed weights on Hugging Face (inclusionAI/Ling-3.0-flash-VL). No first-party USD API pricing published yet.
Model Specs
Released2026-09-04
Parameters124B (5.5B activated)
Context131k
ArchitectureMixture of Experts