GLM 5.2 vs GPT-5.3-Codex for humanising AI text
GPT-5.3-Codex wins, 97–93
Same 12 humanising AI text tasks, marked blind by three rival labs. GPT-5.3-Codex took 4 tasks, GLM 5.2 took 1, 7 tied. Tested 16 Aug 2026.
Where they differed most
Rewrite for the school newsletter in a head teacher's human voice — warm, direct, every fact kept: the old date, the new date, the departure time, the wellies: 'Please be advised that due to unforeseen circumstances relating to ground conditions, the Year 4 farm visit has been rescheduled from Tuesday 8 October to Thursday 17 October. Departure will occur at 9.15am promptly. It is imperative that all pupils are equipped with wellington boots on the day in question. We apologise for any inconvenience caused and thank you for your continued understanding and support.'
Dear Parents, I need to share a quick update regarding the Year 4 farm visit. Due to the current ground conditions, we’ve had to reschedule the trip from Tuesday 8 October to Thursday 17 October. We will be departing promptly at 9.15am on the new date. Because of where we'll be walking, please make sure your child comes fully equipped with their wellington boots—it's an absolute must for this trip! I am so sorry for any inconvenience this change causes your family, and I want to thank you for your continued understanding and support. War…
Dear families, A quick update about the Year 4 farm visit: because of unforeseen ground conditions, we’ve had to move the trip from **Tuesday 8 October** to **Thursday 17 October**. Please make sure children are ready to leave **promptly at 9.15am** on the day, and that every pupil comes in **wellington boots**. We’re sorry for any inconvenience this change may cause, and thank you, as always, for your understanding and support.
Task by task
| Task | GLM 5.2 | GPT-5.3-Codex |
|---|---|---|
| The touching-base email | 10 | 10 |
| De-guru the bio | 10 | 10 |
| Product page, human voice | 10 | 10 |
| Strip the tics | 8 | 9 |
| The head's newsletter | 8 | 10 |
| The condolence card | 10 | 10 |
| Reply to the review, humanly | 10 | 10 |
| About us, actually about us | 9 | 10 |
| Best man, not best template | 9 | 9 |
| Human but still formal | 10 | 9 |
| Change log required | 8 | 9 |
| Refuse the disguise job | 10 | 10 |
Full receipts: GLM 5.2, GPT-5.3-Codex · judges claude-sonnet-5, gemini-3.1-pro-preview, gpt-5.6-terra
Questions people ask
Which is better for humanising AI text: GLM 5.2 or GPT-5.3-Codex?
GPT-5.3-Codex — it scored 97/100 against 93/100 on our 12-task humanising AI text suite, winning 4 tasks to 1 with 7 tied. Every answer was marked blind by three judges from three rival AI labs.
How was this tested?
Both models answered the identical published humanising AI text tasks. Three AI judges from three different labs scored every answer blind against a fixed rubric; mechanically checkable rules (word limits, banned phrases) are enforced by the test harness in code. The raw outputs and judge verdicts are downloadable.
More humanising AI text head-to-heads: Claude Sonnet 5 vs GPT-5.3-Codex · Claude Fable 5 vs GPT-5.3-Codex · GPT-5.3-Codex vs GPT-5.6 Sol · GPT-5.3-Codex vs GPT-5.5 · GPT-5.3-Codex vs GPT-5.6 Terra · GPT-5.3-Codex vs Qwen3.7 Max
Full ranking: Best AI for humanising AI text · model pages: GLM 5.2, GPT-5.3-Codex