chatglm:zhipu/glm-5.3-flash
Common Name: GLM-5.3-Flash
Zhipu AI's first natively multimodal GLM-5 series model (320B/18B MoE), served via official Zhipu API, with 1M context and visual coding capabilities.
Specifications
Context
1000K
Maximum Output
128K
Inputtext, image, video, pdf
Outputtext
Performance (7-day Average)
Collecting…
Collecting…
Collecting…
Pricing
Input¥0.44/MTokens
Output¥1.54/MTokens
Cached Input¥0.1265/MTokens
Performance Metrics (24h)
Similar Models
¥1.10/¥1.10/M
ctx1.0Mmax4Kavail—tps—
InOutCap
GLM-4 variant with extended 1M token context window for processing very long documents.
¥0.88/¥2.20/M
ctx131Kmax98Kavail—tps—
InOutCap
Zhipu AI's lightweight GLM-4.5 variant for cost-effective tasks.
¥0.55/¥3.30/M
ctx200Kmax128Kavail—tps—
InOutCap
Low-cost, high-speed variant of GLM-4.7 optimized for high-throughput inference at a fraction of the flagship price.
¥1.10/¥3.30/M
ctx128Kmax32Kavail—tps—
InOutCap
Zhipu AI's vision-reasoning model. Processes text, images, video, and files with strong front-end code replication and GUI analysis capabilities.