GPT-5.6 Sollatest from openai

People use it to write computer code and solve complex reasoning problems.

openai’s newest large language model, able to hold about 1050k tokens of context (roughly 788k words) in one conversation, working across text and image and file->text. You reach it through an API or through tools built on it. Built for programmers who need help with complex coding tasks. openai describes it as: “GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. source ↗

Our tested score
94/100 · Humanising AI text
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
98/100 · Book writing
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
93/100 · Health questions
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
94/100 · Legal questions
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
96/100 · Emotional support
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
94/100 · Travel planning
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
96/100 · Translation
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
97/100 · Creative writing
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
98/100 · Research
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
93/100 · CV writing
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
95/100 · Presentations
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
97/100 · Vibe coding
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
95/100 · Job applications
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
100/100 · Social posts
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
98/100 · Flashcards
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
93/100 · Coding
18 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
94/100 · Study notes
12 tasks, marked blind by 3 judges · 14 Aug 2026 · receipts
100/100 · Customer replies
12 tasks, marked blind by 3 judges · 14 Aug 2026 · receipts
99/100 · Maths
12 tasks, marked blind by 3 judges · 14 Aug 2026 · receipts
99/100 · Emails
12 tasks, marked blind by 3 judges · 14 Aug 2026 · receipts
100/100 · Extraction
12 tasks, marked blind by 3 judges · 13 Aug 2026 · receipts
95/100 · Summarising
12 tasks, marked blind by 3 judges · 13 Aug 2026 · receipts
95/100 · Essays
12 tasks, marked blind by 3 judges · 13 Aug 2026 · receipts
97/100 · Spreadsheets
12 tasks, marked blind by 3 judges · 13 Aug 2026 · receipts
API price (per 1M in/out)
$2.5 / $15
Battle record
unfought
no battle yet — we don't guess
Forum sentiment (90d)
50%
developer-forum posts: 22 opinions in 104

Compare GPT-5.6 Sol against…

Every claim we hold (3)