GLM-5.3 Flash

Z.ai GLM-5.3-Flash (320B MoE / 18B active) served on GMI Cloud. Natively multimodal with a 1M-token context window — strong coding and agentic performance at flash-tier pricing.

glm-5.3-flash
STABLEGet Started
1,048,576 context
Starting at $0.07/M (50% off) input tokens
Starting at $0.25/M (50% off) output tokens
Streaming
Vision
Tools
Reasoning
JSON Output

Select Provider

All Providers for GLM-5.3 Flash

LLM Gateway routes requests to the best providers that are able to handle your prompt size and parameters.

GMI Cloud
Context: 1.0M50% off
Input
$0.15$0.075
/M tokens
Cached
$0.03$0.015
/M tokens
Output
$0.5$0.25
/M tokens
Get Started