GLM-5.3-FlashX

Z.ai·Pendingtextvisionvideo
Context1,000k
Max output131.072k
Input / 1M$0.37 ↗
Output / 1M$1.25 ↗
Cached in$0.075
Benchmarks0

Benchmarks

Cited public sources · rank is exact-benchmark, apples-to-apples

No cited benchmarks yet.

GLM-5.3-FlashX FAQ

What is GLM-5.3-FlashX?
GLM-5.3-FlashX is a large language model from Z.ai. It accepts text, image and video input.
How much does GLM-5.3-FlashX cost?
GLM-5.3-FlashX costs $0.37 per 1M input tokens and $1.25 per 1M output tokens through the Z.ai API. Cached input costs $0.075 per 1M tokens. Prices are from Z.ai's official pricing as of October 5, 2026.
What is GLM-5.3-FlashX's context window?
GLM-5.3-FlashX has a 1M-token context window and can generate up to 131K output tokens.

Access GLM-5.3-FlashX and every other model through one endpoint with automatic failover: Respan gateway.