Z.ai: GLM 5.3 FlashX.
GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...
View on OpenRouter ↗Benchmarks
How these are aggregated →Profile
Z.ai: GLM 5.3 FlashX is Z.ai's entry at position #18 on the current aggregated leaderboard. Its strongest signals are graduate-level science reasoning, real-world software engineering, unattended agentic work, and raw output speed. Every figure below is aggregated from the public sources listed on the homepage and refreshed on each pipeline run.
The score spread is comparatively complete for this model, but the standing advice still applies: treat gaps of a few points as ties and validate on your own workload before committing.
How this model's position is computed is documented in the methodology; what each benchmark actually measures is in the benchmark guide. See a figure that disagrees with its source? Report it — accepted corrections apply on the next refresh.
Nearby on the leaderboard
- Elo 1638
- Elo 1619
- Elo 1627
- Elo 1608