Grok 4.5latest from x-ai

It helps people write computer code, do knowledge work, and solve science and math problems.

x-ai’s newest large language model, able to hold about 500k tokens of context (roughly 375k words) in one conversation, working across text and image and file->text. You reach it through an API or through tools built on it. Built for coders, knowledge workers, and people working in STEM fields. x-ai describes it as: “Grok 4.5 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM. source ↗

Our tested score
90/100 · Humanising AI text
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
90/100 · Book writing
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
95/100 · Health questions
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
86/100 · Legal questions
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
90/100 · Emotional support
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
83/100 · Travel planning
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
83/100 · Translation
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
93/100 · Creative writing
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
93/100 · Research
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
88/100 · CV writing
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
91/100 · Presentations
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
90/100 · Vibe coding
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
88/100 · Job applications
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
93/100 · Social posts
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
94/100 · Flashcards
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
88/100 · Coding
18 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
88/100 · Study notes
12 tasks, marked blind by 3 judges · 14 Aug 2026 · receipts
98/100 · Customer replies
12 tasks, marked blind by 3 judges · 14 Aug 2026 · receipts
98/100 · Maths
12 tasks, marked blind by 3 judges · 14 Aug 2026 · receipts
93/100 · Emails
12 tasks, marked blind by 3 judges · 14 Aug 2026 · receipts
100/100 · Extraction
12 tasks, marked blind by 3 judges · 13 Aug 2026 · receipts
92/100 · Summarising
12 tasks, marked blind by 3 judges · 13 Aug 2026 · receipts
83/100 · Essays
12 tasks, marked blind by 3 judges · 13 Aug 2026 · receipts
94/100 · Spreadsheets
12 tasks, marked blind by 3 judges · 13 Aug 2026 · receipts
92/100 · Best-value API
30 tasks, marked blind by 3 judges · 11 Aug 2026 · receipts
API price (per 1M in/out)
$2 / $6
Community Elo (text)
1451
Battle record
2W – 2L
in our published battles
Forum sentiment (90d)
100%
developer-forum posts: 4 opinions in 151
Battles fought

Compare Grok 4.5 against…

Every claim we hold (7)