fireworks/fast:zhipu/glm-5.3

Common Name: GLM-5.3 Fast

Fireworks
-50%On SaleSupportedTool InvocationSupportedReasoning
CompareTry in Chat

Z.ai's accelerated GLM-5.3 reasoning model for low-latency agentic workloads.

Specifications

Context
1048.6K
Maximum Output
131.1K
Inputtext
Outputtext

Performance (7-day Average)

Collecting…
Collecting…
Collecting…

Availability Trend (24h)

Pricing

Input$1.05/MTokens
Cached Input$0.195/MTokens
Output$3.30/MTokens

Performance Metrics (24h)

Similar Models

$1.50/$7.50/M-50%
ctx1.0Mmax131Kavail—tps—
InOutCap

Fireworks' specialized model built on Kimi K3 with shorter reasoning traces for efficient agentic workloads.

$0.70/$2.20/M-50%
ctx1.0Mmax131Kavail—tps—
InOutCap

Z.ai's flagship reasoning model with a 1M-token context window and controllable thinking length.

$1.00/$3.00/M-50%
ctx1.0Mmax131Kavail—tps—
InOutCap

Alibaba's multimodal reasoning model for general-purpose and agentic workloads.

$1.50/$7.50/M-50%
ctx1.0Mmax131Kavail—tps—
InOutCap

Kimi's flagship multimodal reasoning model for long-context knowledge work and agentic workflows.