GLM 5.2latest from z-ai

It processes large amounts of text to help users write software and solve complex problems.

z-ai’s newest large language model, able to hold about 1049k tokens of context (roughly 786k words) in one conversation, working across text->text. You reach it through an API or through tools built on it. Built for software engineers working on large projects. z-ai describes it as: “GLM 5.2 is a large-scale reasoning model from Z.ai. source ↗

Our tested score
88/100 · Meetings & notes
12 tasks, marked blind by 3 judges · 28 Aug 2026 · receipts
81/100 · Bookkeeping & accounts
12 tasks, marked blind by 3 judges · 18 Aug 2026 · receipts
78/100 · Property & lettings
12 tasks, marked blind by 3 judges · 18 Aug 2026 · receipts
83/100 · HR & employment
12 tasks, marked blind by 3 judges · 18 Aug 2026 · receipts
82/100 · Workflow automation
12 tasks, marked blind by 3 judges · 18 Aug 2026 · receipts
78/100 · Research agents
12 tasks, marked blind by 3 judges · 18 Aug 2026 · receipts
87/100 · Code review
12 tasks, marked blind by 3 judges · 18 Aug 2026 · receipts
93/100 · Humanising AI text
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
90/100 · Book writing
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
89/100 · Health questions
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
86/100 · Legal questions
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
91/100 · Emotional support
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
76/100 · Travel planning
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
86/100 · Translation
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
86/100 · Creative writing
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
86/100 · Research
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
85/100 · CV writing
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
87/100 · Presentations
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
81/100 · Vibe coding
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
95/100 · Job applications
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
93/100 · Social posts
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
93/100 · Flashcards
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
92/100 · Coding
18 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
89/100 · Study notes
12 tasks, marked blind by 3 judges · 14 Aug 2026 · receipts
93/100 · Customer replies
12 tasks, marked blind by 3 judges · 14 Aug 2026 · receipts
97/100 · Maths
12 tasks, marked blind by 3 judges · 14 Aug 2026 · receipts
93/100 · Emails
12 tasks, marked blind by 3 judges · 14 Aug 2026 · receipts
96/100 · Extraction
12 tasks, marked blind by 3 judges · 13 Aug 2026 · receipts
91/100 · Summarising
12 tasks, marked blind by 3 judges · 13 Aug 2026 · receipts
79/100 · Essays
12 tasks, marked blind by 3 judges · 13 Aug 2026 · receipts
96/100 · Spreadsheets
12 tasks, marked blind by 3 judges · 13 Aug 2026 · receipts
92/100 · Best-value API
30 tasks, marked blind by 3 judges · 11 Aug 2026 · receipts
API price (per 1M in/out)
$0.6496 / $2.0416
Battle record
2W – 2L
in our published battles
Forum sentiment — 90d to 12 Aug 2026
67%
developer-forum posts: 9 opinions in 68
stale — measured 41 days ago, not re-swept since
Battles fought

Compare GLM 5.2 against…

Every claim we hold (4)