NanoGPT
GLM 5.3 Flash
Native multimodal GLM model for efficient coding and long-horizon agent tasks
z-ai/glm-5.3-flashContext
1M
Max output
131.1K
Input $/M
$0.075
Output $/M
$0.25
Cache read $/M
$0.015
Cache write $/M
—
Against the catalogue
Context window1M
Larger than 95% of the 7,784 models listed
Input price$0.075
Cheaper than 85% of priced models
Capabilities
Reasoningsupported
Tool callingsupported
Structured outputsupported
Attachmentssupported
Temperaturesupported
Open weightssupported
Interleaved thinkingnot supported
Modalities
Input
TextImageVideo
Output
Text
Reasoning controls
effortlow, high, maxPricing detail
| Rate | Base $/M |
|---|---|
| Input | $0.075 |
| Output | $0.25 |
| Cache read | $0.015 |