The newest GLM Flash model is GLM-5.3-Flash, released on August 26, 2026. It has been 45 days since then. The line has only 1 gap between releases so far (219 days), too few to know its usual rhythm.
Where the line stands
How long it has been since the last release against the line's usual gap and its recent ones — the rhythm, not a date.
Get new AI models by email
Every Monday, only when something shipped: the AI models that came out, what each one changed, and which lines are due next.
We email you a link to confirm first · unsubscribe anytime · read past issues · about the newsletter
Every GLM Flash release
Newest first — what each version changed versus the one before, in Z.ai's own numbers.
GLM-5.3-Flash is a 320B-parameter open model (18B active) with vision and a 1M-token context window that Z.ai compares with GLM-5.2: it solves 63.4% of DeepSWE v1.1 coding tasks versus 46.2%, at about a tenth of the API price.
Compared with GLM-5.2 · comparison chosen by Z.ai
Solves over 6 in 10 long real-world coding tasks
DeepSWE v1.1 — long software-engineering tasks in real code repositories
63.4%
▲ GLM-5.2 scored 46.2% · +17.2 points
Scores 8 points higher than GLM-5.2 on an independent intelligence test
Artificial Analysis Intelligence Index v4.3.2 — one score averaging 10 hard tests
42
▲ GLM-5.2 scores 34 · +8 points
Costs developers about a tenth of GLM-5.2
Price per 1M tokens — what developers pay, input / output
$0.15 / $0.50
▲ GLM-5.2 costs $1.40 / $4.40