glm-4-flashx

Common Name: GLM-4 FlashX

ChatGLM
Released on Feb 17, 2025 12:00 AMSupportedTool Invocation
CompareTry in Chat

Zhipu AI's fastest GLM-4 variant optimized for high-throughput inference.

Specifications

Context
128K
Maximum Output
16.4K
Inputtext
Outputtext

Performance (7-day Average)

Collecting…
Collecting…
Collecting…

Pricing

Input¥0.11/MTokens
Output¥0.11/MTokens
Batch Input¥0.055/MTokens
Batch Output¥0.055/MTokens

Availability Trend (24h)

Performance Metrics (24h)

Similar Models

¥0.165/¥1.65/M
ctx128Kmax32Kavailtps
InOutCap

Lightweight, high-speed variant of GLM-4.6V with multimodal tool calling and long-context visual reasoning.

¥0.22/¥0.22/M
ctx131Kmaxavailtps
InOut

Zhipu AI's 0.9B document OCR model (94.62 on OmniDocBench V1.5). Converts images and PDFs to Markdown while preserving table, formula, and layout structure. Served through the layout parsing endpoint, not chat completions; a single call accepts one image (≤10MB) or PDF (≤50MB, ≤100 pages).

¥2.20/¥6.60/M
ctx64Kmax16Kavailtps
InOutCap

Zhipu AI's multimodal model with vision capabilities. Processes text, images, video, and files for analysis tasks.

Free/Free
ctx131Kmax98Kavailtps
InOutCap

Fast, cost-efficient version of GLM-4.5. Optimized for high-throughput applications.