Claude Sonnet 5 vs GPT-5.3-Codex for humanising AI text

The verdict

Effectively a dead heat: 9796

Same 12 humanising AI text tasks, marked blind by three rival labs. GPT-5.3-Codex took 1 task, Claude Sonnet 5 took 0, 11 tied. Tested 16 Aug 2026.

Where they differed most

The task — The head's newsletter

Rewrite for the school newsletter in a head teacher's human voice — warm, direct, every fact kept: the old date, the new date, the departure time, the wellies: 'Please be advised that due to unforeseen circumstances relating to ground conditions, the Year 4 farm visit has been rescheduled from Tuesday 8 October to Thursday 17 October. Departure will occur at 9.15am promptly. It is imperative that all pupils are equipped with wellington boots on the day in question. We apologise for any inconvenience caused and thank you for your continued understanding and support.'

A quick change of plan for our Year 4 farm visit!

Because of the ground conditions at the farm, we've had to move the trip from Tuesday 8 October to Thursday 17 October. We'll now be setting off at 9.15am promptly, so please make sure your child arrives in good time that morning.

One important reminder: wellies are a must! Farm visits are always more fun (and less muddy) with the right footwear.

Sorry for any inconvenience this change causes with your plans — thank you, as ever, for your understanding and support.
Dear families,

A quick update about the Year 4 farm visit: because of unforeseen ground conditions, we’ve had to move the trip from **Tuesday 8 October** to **Thursday 17 October**.  
Please make sure children are ready to leave **promptly at 9.15am** on the day, and that every pupil comes in **wellington boots**.

We’re sorry for any inconvenience this change may cause, and thank you, as always, for your understanding and support.

Task by task

TaskClaude Sonnet 5GPT-5.3-Codex
The touching-base email1010
De-guru the bio1010
Product page, human voice1010
Strip the tics99
The head's newsletter910
The condolence card1010
Reply to the review, humanly1010
About us, actually about us1010
Best man, not best template99
Human but still formal99
Change log required99
Refuse the disguise job1010

Full receipts: Claude Sonnet 5, GPT-5.3-Codex · judges gemini-3.1-pro-preview, gpt-5.6-terra, grok-4.5

Questions people ask

Which is better for humanising AI text: Claude Sonnet 5 or GPT-5.3-Codex?

Effectively a dead heat: GPT-5.3-Codex edged it 97/100 to 96/100 on our humanising AI text suite — too close to matter, so pick on price or the product you already use.

How was this tested?

Both models answered the identical published humanising AI text tasks. Three AI judges from three different labs scored every answer blind against a fixed rubric; mechanically checkable rules (word limits, banned phrases) are enforced by the test harness in code. The raw outputs and judge verdicts are downloadable.

More humanising AI text head-to-heads: Claude Fable 5 vs GPT-5.3-Codex · GPT-5.3-Codex vs GPT-5.6 Sol · GPT-5.3-Codex vs GPT-5.5 · GLM 5.2 vs GPT-5.3-Codex · GPT-5.3-Codex vs GPT-5.6 Terra · GPT-5.3-Codex vs Qwen3.7 Max

Full ranking: Best AI for humanising AI text · model pages: Claude Sonnet 5, GPT-5.3-Codex