Context
1.0M tokens (~524 books)
Input $/1M
$0.37
Output $/1M
$1.25
Type
multimodal
License
Proprietary
Benchmarks
0 tested
Data updated today
About
GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...
No benchmark data available yet.
Links
Research
Documentation
Community
BenchGecko API
glm-5-3-flashx
Specifications
- Typemultimodal
- Context1.0M tokens (~524 books)
- ReleasedSep 2026
- LicenseProprietary
- StatusActive
- Cost / Message~$0.002
Available On
Learn More
Share & Export
Frequently Asked Questions
GLM 5.3 FlashX is a proprietary multimodal AI model by z-ai, released in September 2026. Context window: 1M tokens.