fireworks:zhipu/glm-5.3-flash

Common Name: GLM-5.3-Flash

Fireworks
-50%On SaleReleased on Aug 26 12:00 AMSupportedTool InvocationSupportedReasoning
CompareTry in Chat

Z.ai's multimodal GLM-5.3 model for fast reasoning and visual coding workloads.

Specifications

Context
1048.6K
Maximum Output
131.1K
Inputtext, image, video, pdf
Outputtext

Performance (7-day Average)

Collecting…
Collecting…
Collecting…

Availability Trend (24h)

Pricing

Input$0.075/MTokens
Cached Input$0.015/MTokens
Output$0.25/MTokens

Performance Metrics (24h)

Similar Models

$0.11/$0.33/M-50%
ctx1.0Mmax384Kavail—tps—
InOutCap

DeepSeek V4.1 Flash with native multimodal understanding and efficient reasoning for agentic workloads.

$0.15/$0.60/M-50%
ctx512Kmax512Kavail—tps—
InOutCap

MiniMax's reasoning model for long-context coding and agentic tasks.

$0.075/$0.30/M-50%
ctx131Kmax33Kavail—tps—
InOutCap

OpenAI's open-weight 120B model for production, agentic tasks, and high-reasoning use cases.

$1.50/$7.50/M-50%
ctx1.0Mmax131Kavail—tps—
InOutCap

Fireworks' specialized model built on Kimi K3 with shorter reasoning traces for efficient agentic workloads.