Vancine
GLM-5.3-Flash
Native multimodal GLM model for efficient coding and long-horizon agent tasks
glm-5.3-flashContext
1M
Max output
131.1K
Input $/M
$0.06
Output $/M
$0.20
Cache read $/M
$0.012
Cache write $/M
—
Against the catalogue
Context window1M
Larger than 82% of the 7,784 models listed
Input price$0.06
Cheaper than 86% of priced models
Capabilities
Reasoningsupported
Tool callingsupported
Structured outputsupported
Attachmentssupported
Temperaturesupported
Open weightssupported
Interleaved thinkingnot supported
Modalities
Input
TextImageVideoPDF
Output
Text
Reasoning controls
effortlow, high, maxPricing detail
| Rate | Base $/M |
|---|---|
| Input | $0.06 |
| Output | $0.20 |
| Cache read | $0.012 |