Claude Sonnet 5latest from anthropic

People use it to write code, complete professional tasks, and solve problems with adaptive thinking.

anthropic’s newest large language model, able to hold about 1000k tokens of context (roughly 750k words) in one conversation, working across text and image and file->text. You reach it through an API or through tools built on it. Built for professionals and developers who write code and do professional work. anthropic describes it as: “Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. source ↗

Our tested score
96/100 · Humanising AI text
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
89/100 · Book writing
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
88/100 · Health questions
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
88/100 · Legal questions
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
98/100 · Emotional support
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
88/100 · Travel planning
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
91/100 · Translation
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
93/100 · Creative writing
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
84/100 · Research
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
83/100 · CV writing
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
88/100 · Presentations
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
76/100 · Vibe coding
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
90/100 · Job applications
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
86/100 · Social posts
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
86/100 · Flashcards
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
83/100 · Coding
18 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
81/100 · Study notes
12 tasks, marked blind by 3 judges · 14 Aug 2026 · receipts
93/100 · Customer replies
12 tasks, marked blind by 3 judges · 14 Aug 2026 · receipts
90/100 · Maths
12 tasks, marked blind by 3 judges · 14 Aug 2026 · receipts
92/100 · Emails
12 tasks, marked blind by 3 judges · 14 Aug 2026 · receipts
67/100 · Extraction
12 tasks, marked blind by 3 judges · 13 Aug 2026 · receipts
93/100 · Summarising
12 tasks, marked blind by 3 judges · 13 Aug 2026 · receipts
85/100 · Essays
12 tasks, marked blind by 3 judges · 13 Aug 2026 · receipts
78/100 · Spreadsheets
12 tasks, marked blind by 3 judges · 13 Aug 2026 · receipts
API price (per 1M in/out)
$2 / $10
Battle record
unfought
no battle yet — we don't guess
Forum sentiment (90d)
40%
developer-forum posts: 5 opinions in 93

Compare Claude Sonnet 5 against…

Every claim we hold (3)