z-ai/glm-5.3-flashGLM-5.3-Flash is Z.ai (Zhipu AI)’s open-weights AI model, released Aug 20, 2026. It costs $0.15/M input and $0.50/M output tokens, and handles a 1.3M-token context window. Its strongest use case is coding, where it ranks in the 83rd percentile of the models Model Beat tracks.
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks.
Scores on standardized evaluations. Higher is better — and rank shows where GLM-5.3-Flash lands among all models tracked on Model Beat.
Graduate-level scientific reasoning
Frontier of human expert knowledge
Factual accuracy & hallucination
Scientific research coding
Human-rated web development
Olympiad-qualifier math
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks.
GLM-5.3-Flash is an AI model developed by Z.ai (Zhipu AI), released Aug 20, 2026.
GLM-5.3-Flash's list price from Z.ai (Zhipu AI) is $0.15 per million input tokens and $0.50 per million output tokens. The cheapest credible third-party provider on OpenRouter (Wafer) serves it at $0.10/$0.35 per 1M.
GLM-5.3-Flash supports a context window of up to 1.3M tokens.
On standardized evaluations tracked by Epoch AI, GLM-5.3-Flash scores 1604 on WebDev Arena, 39.9% on Humanity's Last Exam, 51.6% on SciCode.
GLM-5.3-Flash is released with open weights (Open weights (unrestricted)), so the model can be downloaded and self-hosted.
Benchmark data from Epoch AI (CC BY) and Artificial Analysis; pricing, specs & descriptions from OpenRouter. Data updated Sep 10, 2026.