LLM Reference
OpenRouter

Using Ling-3.0-flash-VL on OpenRouter

Implementation guide · Ling 3.0 · InclusionAI

ServerlessOpen Source

OpenRouter exposes Ling-3.0-flash-VL through model ID inclusionai/ling-3.0-flash-vl. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-09-08. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at OpenRouter and generate an API key.
  2. 2
    Use the OpenRouter SDK or REST API to call inclusionai/ling-3.0-flash-vl — see the documentation for request format.
  3. 3
    You'll be billed $0.06/1M input, $0.18/1M output tokens. See full pricing.

Code Examples

See OpenRouter documentation for integration details.

Pricing on OpenRouter

TypePrice (per 1M)
Input tokens$0.06
Output tokens$0.18

Capabilities

VisionMultimodalReasoningJSON / Tool useStructured Outputs

About Ling-3.0-flash-VL

Ling-3.0-flash-VL is InclusionAI's vision-language fine-tune of Ling-3.0-flash (124B total / 5.5B activated MoE) for image and video understanding, reasoning, function calling, and agentic tool use. Native config context is 131,072 tokens; the Hugging Face model card documents YaRN extension up to 1M tokens. MIT-licensed weights on Hugging Face (inclusionAI/Ling-3.0-flash-VL). No first-party USD API pricing published yet.

Model Specs

Released2026-09-04
Parameters124B (5.5B activated)
Context131k
ArchitectureMixture of Experts

Provider

OpenRouter
OpenRouter

OpenRouter, Inc.

New York, NY, USA