NanoGPT
Ling 3.0 Flash VL
Ling 3.0 Flash VL is inclusionAI's native multimodal Mixture-of-Experts model with 124B total parameters and 5.5B active parameters per token. It combines image and video understanding with reasoning and tool use for document analysis, charts, visual verification, and interface-based agent tasks. Thinking is enabled by default and can be turned off in settings.
inclusionai/ling-3.0-flash-vlContext
262.1K
Max output
32.8K
Input $/M
$0.06
Output $/M
$0.18
Cache read $/M
$0.012
Cache write $/M
—
Against the catalogue
Context window262.1K
Larger than 57% of the 7,784 models listed
Input price$0.06
Cheaper than 86% of priced models
Capabilities
Reasoningsupported
Tool callingsupported
Structured outputnot supported
Attachmentssupported
Temperaturenot supported
Open weightssupported
Interleaved thinkingnot supported
Modalities
Input
TextImageVideo
Output
Text
Reasoning controls
effortnone, low, medium, highPricing detail
| Rate | Base $/M |
|---|---|
| Input | $0.06 |
| Output | $0.18 |
| Cache read | $0.012 |