GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
Z.ai: GLM 5.3 Flash — Z.ai. Context window 1,048,576 tokens. API pricing (per 1M tokens): Input $0.15, Output $0.5. License: Proprietary. LMArena 1,470. GPQA Diamond 90.2%.
| Provider | Z.ai |
|---|---|
| API model ID | z-ai/glm-5.3-flash |
| Tier | Frontier |
| Context window | 1,048,576 tokens |
| Max output | 943,717 tokens |
| License | Proprietary |
| API pricing (per 1M tokens) · Input | $0.15 |
| API pricing (per 1M tokens) · Output | $0.5 |
| LMArena | 1,470 (glm-5.3-flash) |
| GPQA Diamond | 90.2% |
LMArena (CC-BY 4.0, best reasoning setting, 2026-10-02) · GPQA Diamond: Epoch AI (CC-BY 4.0)
Compare in the full catalog