glm-4-flashx
Common Name: GLM-4 FlashX
Zhipu AI's fastest GLM-4 variant optimized for high-throughput inference.
Specifications
Context
128K
Maximum Output
16.4K
Inputtext
Outputtext
Performance (7-day Average)
Collecting…
Collecting…
Collecting…
Pricing
Input¥0.11/MTokens
Output¥0.11/MTokens
Batch Input¥0.055/MTokens
Batch Output¥0.055/MTokens
Availability Trend (24h)
Performance Metrics (24h)
Similar Models
¥0.165/¥1.65/M
ctx128Kmax32Kavail—tps—
InOutCap
Lightweight, high-speed variant of GLM-4.6V with multimodal tool calling and long-context visual reasoning.
¥0.22/¥0.22/M
ctx131Kmax—avail—tps—
InOut
Zhipu AI's 0.9B document OCR model (94.62 on OmniDocBench V1.5). Converts images and PDFs to Markdown while preserving table, formula, and layout structure. Served through the layout parsing endpoint, not chat completions; a single call accepts one image (≤10MB) or PDF (≤50MB, ≤100 pages).
¥2.20/¥6.60/M
ctx64Kmax16Kavail—tps—
InOutCap
Zhipu AI's multimodal model with vision capabilities. Processes text, images, video, and files for analysis tasks.
Free/Free
ctx131Kmax98Kavail—tps—
InOutCap
Fast, cost-efficient version of GLM-4.5. Optimized for high-throughput applications.