Gemini 3.1 Pro Preview

It helps software engineers build applications and manage complex workflows.

google’s large language model, able to hold about 1049k tokens of context (roughly 786k words) in one conversation, working across text and image and file and audio and video->text. You reach it through an API or through tools built on it. Built for software engineers. google describes it as: “Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. source ↗

Our tested score
87/100 · Humanising AI text
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
87/100 · Book writing
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
89/100 · Health questions
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
85/100 · Legal questions
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
91/100 · Emotional support
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
79/100 · Travel planning
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
83/100 · Translation
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
95/100 · Creative writing
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
85/100 · Research
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
83/100 · CV writing
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
83/100 · Presentations
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
81/100 · Vibe coding
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
90/100 · Job applications
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
88/100 · Social posts
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
95/100 · Flashcards
12 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
88/100 · Coding
18 tasks, marked blind by 3 judges · 16 Aug 2026 · receipts
89/100 · Study notes
12 tasks, marked blind by 3 judges · 14 Aug 2026 · receipts
97/100 · Customer replies
12 tasks, marked blind by 3 judges · 14 Aug 2026 · receipts
96/100 · Maths
12 tasks, marked blind by 3 judges · 14 Aug 2026 · receipts
88/100 · Emails
12 tasks, marked blind by 3 judges · 14 Aug 2026 · receipts
100/100 · Extraction
12 tasks, marked blind by 3 judges · 13 Aug 2026 · receipts
90/100 · Summarising
12 tasks, marked blind by 3 judges · 13 Aug 2026 · receipts
83/100 · Essays
12 tasks, marked blind by 3 judges · 13 Aug 2026 · receipts
93/100 · Spreadsheets
12 tasks, marked blind by 3 judges · 13 Aug 2026 · receipts
86/100 · Best-value API
30 tasks, marked blind by 3 judges · 11 Aug 2026 · receipts
API price (per 1M in/out)
$2 / $12
Community Elo (text)
1480
Battle record
unfought
no battle yet — we don't guess
Forum sentiment (90d)
9%
developer-forum posts: 11 opinions in 87

Compare Gemini 3.1 Pro Preview against…

Every claim we hold (6)