glm-4.6v-flashx

Common Name: GLM-4.6V FlashX

ChatGLM
Released on Oct 8, 2025 12:00 AMKnowledge Cutoff Apr 1, 2025 12:00 AMSupportedTool InvocationSupportedReasoning
CompareTry in Chat

Lightweight, high-speed variant of GLM-4.6V with multimodal tool calling and long-context visual reasoning.

Specifications

Context
128K
Maximum Output
32K
Inputtext, image, video, pdf
Outputtext

Performance (7-day Average)

Collecting…
Collecting…
Collecting…

Pricing

< 32K
Input¥0.165/MTokens
Output¥1.65/MTokens
Cached Input¥0.033/MTokens
32K-128K
Input¥0.33/MTokens
Output¥3.30/MTokens
Cached Input¥0.033/MTokens

Availability Trend (24h)

Performance Metrics (24h)

Similar Models

¥0.22/¥0.22/M
ctx131Kmaxavailtps
InOut

Zhipu AI's 0.9B document OCR model (94.62 on OmniDocBench V1.5). Converts images and PDFs to Markdown while preserving table, formula, and layout structure. Served through the layout parsing endpoint, not chat completions; a single call accepts one image (≤10MB) or PDF (≤50MB, ≤100 pages).

¥0.11/¥0.11/M
ctx128Kmax16Kavailtps
InOutCap

Zhipu AI's fastest GLM-4 variant optimized for high-throughput inference.

¥2.20/¥6.60/M
ctx64Kmax16Kavailtps
InOutCap

Zhipu AI's multimodal model with vision capabilities. Processes text, images, video, and files for analysis tasks.

Free/Free
ctx131Kmax98Kavailtps
InOutCap

Fast, cost-efficient version of GLM-4.5. Optimized for high-throughput applications.