glm-5.3-flash

Common Name: GLM-5.3-Flash

ChatGLM
-45%On SaleReleased on Aug 26 12:00 AMSupportedTool InvocationSupportedReasoning
CompareTry in Chat

Zhipu AI's first natively multimodal GLM-5 series model (320B/18B MoE), served via official Zhipu API, with 1M context and visual coding capabilities.

Specifications

Context
1000K
Maximum Output
128K
Inputtext, image, video, pdf
Outputtext

Performance (7-day Average)

Collecting…
Collecting…
Collecting…

Pricing

Input¥0.44/MTokens
Output¥1.54/MTokens
Cached Input¥0.1265/MTokens

Performance Metrics (24h)

Similar Models

¥1.10/¥1.10/M
ctx1.0Mmax4Kavailtps
InOutCap

GLM-4 variant with extended 1M token context window for processing very long documents.

¥0.88/¥2.20/M
ctx131Kmax98Kavailtps
InOutCap

Zhipu AI's lightweight GLM-4.5 variant for cost-effective tasks.

¥0.55/¥3.30/M
ctx200Kmax128Kavailtps
InOutCap

Low-cost, high-speed variant of GLM-4.7 optimized for high-throughput inference at a fraction of the flagship price.

¥1.10/¥3.30/M
ctx128Kmax32Kavailtps
InOutCap

Zhipu AI's vision-reasoning model. Processes text, images, video, and files with strong front-end code replication and GUI analysis capabilities.