RunInfra
GLM-5.3-Flash
Native multimodal GLM model for efficient coding and long-horizon agent tasks
zai-org/GLM-5.3-FlashContext
1M
Max output
32.8K
Input $/M
$0.10
Output $/M
$0.40
Cache read $/M
$0.01
Cache write $/M
—
Against the catalogue
Context window1M
Larger than 95% of the 7,784 models listed
Input price$0.10
Cheaper than 80% of priced models
Capabilities
Reasoningsupported
Tool callingsupported
Structured outputsupported
Attachmentssupported
Temperaturesupported
Open weightssupported
Interleaved thinkingsupported
Modalities
Input
TextImage
Output
Text
Reasoning controls
effortlow, high, maxPricing detail
| Rate | Base $/M |
|---|---|
| Input | $0.10 |
| Output | $0.40 |
| Cache read | $0.01 |