GPT-5.3-Codex

It writes software code and assists with software engineering tasks.

openai’s large language model, able to hold about 400k tokens of context (roughly 300k words) in one conversation, working across text and image and file->text. You reach it through an API or through tools built on it. Built for software engineers and developers. openai describes it as: “It achieves state-of-the-art results... source ↗

Our tested score
97/100 · Humanising AI text
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
95/100 · Book writing
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
94/100 · Health questions
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
93/100 · Legal questions
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
95/100 · Emotional support
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
88/100 · Travel planning
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
93/100 · Translation
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
94/100 · Creative writing
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
95/100 · Research
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
93/100 · CV writing
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
97/100 · Presentations
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
98/100 · Vibe coding
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
95/100 · Job applications
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
99/100 · Social posts
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
97/100 · Flashcards
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
92/100 · Coding
18 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
88/100 · Study notes
12 tasks, marked blind by 3 judges · 14 Aug 2026 · receipts
98/100 · Customer replies
12 tasks, marked blind by 3 judges · 14 Aug 2026 · receipts
98/100 · Maths
12 tasks, marked blind by 3 judges · 14 Aug 2026 · receipts
99/100 · Emails
12 tasks, marked blind by 3 judges · 14 Aug 2026 · receipts
100/100 · Extraction
12 tasks, marked blind by 3 judges · 13 Aug 2026 · receipts
96/100 · Summarising
12 tasks, marked blind by 3 judges · 13 Aug 2026 · receipts
88/100 · Essays
12 tasks, marked blind by 3 judges · 13 Aug 2026 · receipts
98/100 · Spreadsheets
12 tasks, marked blind by 3 judges · 13 Aug 2026 · receipts
API price (per 1M in/out)
$1.75 / $14
Battle record
unfought
no battle yet — we don't guess
Forum sentiment (90d)
33%
developer-forum posts: 6 opinions in 105

Compare GPT-5.3-Codex against…

Every claim we hold (3)