Z.ai: GLM 5.3 Flash

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

Z.ai: GLM 5.3 Flash — Z.ai. Context window 1,048,576 tokens. API pricing (per 1M tokens): Input $0.15, Output $0.5. License: Proprietary. LMArena 1,470. GPQA Diamond 90.2%.

Benchmarks · API pricing (per 1M tokens)

ProviderZ.ai
API model IDz-ai/glm-5.3-flash
TierFrontier
Context window1,048,576 tokens
Max output943,717 tokens
LicenseProprietary
API pricing (per 1M tokens) · Input$0.15
API pricing (per 1M tokens) · Output$0.5
LMArena1,470 (glm-5.3-flash)
GPQA Diamond90.2%

LMArena (CC-BY 4.0, best reasoning setting, 2026-10-02) · GPQA Diamond: Epoch AI (CC-BY 4.0)

Compare in the full catalog