NanoGPTLing 3.0 Flash VL

NanoGPT

Ling 3.0 Flash VL

Ling 3.0 Flash VL is inclusionAI's native multimodal Mixture-of-Experts model with 124B total parameters and 5.5B active parameters per token. It combines image and video understanding with reasoning and tool use for document analysis, charts, visual verification, and interface-based agent tasks. Thinking is enabled by default and can be turned off in settings.

inclusionai/ling-3.0-flash-vl
Context
262.1K
Max output
32.8K
Input $/M
$0.06
Output $/M
$0.18
Cache read $/M
$0.012
Cache write $/M

Against the catalogue

Context window262.1K

Larger than 57% of the 7,784 models listed

Input price$0.06

Cheaper than 86% of priced models

Capabilities

Reasoningsupported
Tool callingsupported
Structured outputnot supported
Attachmentssupported
Temperaturenot supported
Open weightssupported
Interleaved thinkingnot supported

Modalities

Input
TextImageVideo
Output
Text

Reasoning controls

effortnone, low, medium, high

Pricing detail

RateBase $/M
Input$0.06
Output$0.18
Cache read$0.012