Parameters
Undisclosed
Context Length
1.05M
Modality
Multimodal
Architecture
Undisclosed
License
Proprietary
Release Date
18 Sept 2026
Knowledge Cutoff
-
API Pricing (per 1M)
Input: $0.37 · Output: $1.25
No evaluation benchmarks for GLM-5.3 FlashX available.
Overall Rank
-
Coding Rank
-
GLM-5.3 FlashX is a high-speed variant of Z.ai's GLM-5.3 Flash native multimodal architecture, reaching inference speeds up to 200 tokens per second. It leverages hybrid sparse and linear attention to deliver fast coding and agent execution over long contexts.
Architecture specifications are undisclosed for proprietary models.
General Language Models from Z.ai
Assistant
Online