Context
1.0M tokens (~524 books)
Input $/1M
$0.37
Output $/1M
$1.25
Type
multimodal
License
Proprietary
Benchmarks
0 tested
Data as of
About
GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...
No benchmark data available yet.
Recently Happened
GLM 5.3 FlashX added
Sep 19, 2026
Links
Research
Documentation
Community
BenchGecko API
glm-5-3-flashx
Specifications
- Typemultimodal
- Context1.0M tokens (~524 books)
- ReleasedSep 2026
- LicenseProprietary
- StatusActive
- Cost / Message~$0.002
Available On
Learn More
Share & Export
Frequently Asked Questions
GLM 5.3 FlashX is a proprietary multimodal AI model by z-ai, released in September 2026. Context window: 1M tokens.
Related Models
Qwen3 Coder Next FP8 · AlibabaQwen2.5 Coder 32B Instruct AWQ · AlibabaSpeaker Diarization Community 1 · PyannoteGemma 3 12B (free) · Google DeepMindQwen3 Next 80B A3B Instruct (free) · Alibaba QwenBenchmarks
Key facts · as of 2026-10-05
- GLM 5.3 FlashX by z-ai. Not enough public benchmark scores to rank yet.
- List price $0.37 input · $1.25 output per 1M tokens (as of 2026-10-05).
- Sold by 1 provider (as of 2026-10-05): Z.AI (fp8) $0.37 in / $1.25 out.
- Gecko Tests: Who Are You B (Knows who made it) · Tokenizer Tax D (70% more tokens outside English).
How to cite · data as of 2026-10-05
GLM 5.3 FlashX · benchmarks, pricing and providers. BenchGecko, data as of 2026-10-05. https://benchgecko.ai/model/glm-5-3-flashx
Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP