Claude Opus 4.8 vs GPT-5.6 Sol for presentations and slides
GPT-5.6 Sol wins, 95–92
Same 12 presentations and slides tasks, marked blind by three rival labs. GPT-5.6 Sol took 2 tasks, Claude Opus 4.8 took 1, 9 tied. Tested 16 Aug 2026.
Where they differed most
This slide is for a 5-minute update to the senior team. Cut it to the bullets that help them DECIDE something — max 4 bullets, nothing else added. Slide 'Office move update': - Project kicked off on 4 June as agreed - Weekly meetings have been held every Tuesday - The project team has worked very hard - Three sites shortlisted, all within 2 miles of the station - Broadband installation lead time is 6 weeks, which is now the critical path - We looked at many options for furniture - Costs are tracking 8% under the £90k budget - The tone of meetings has been very positive - Decision needed by 30 September or we lose the preferred site.
Here are the 4 decision-relevant bullets: - Three sites shortlisted, all within 2 miles of the station - Broadband install has a 6-week lead time — now the critical path - Costs tracking 8% under the £90k budget - Decision needed by 30 September or we lose the preferred site
- Three sites shortlisted, all within 2 miles of the station - Decision needed by 30 September or we lose the preferred site - Broadband installation takes 6 weeks and is now the critical path - Costs are tracking 8% under the £90k budget
Task by task
| Task | Claude Opus 4.8 | GPT-5.6 Sol |
|---|---|---|
| Deck from a standing start | 9 | 9 |
| Speaker notes, not an essay | 10 | 10 |
| One slide, not a memo | 9 | 9 |
| The first 30 seconds | 9 | 10 |
| Slide on a diet | 5 | 9 |
| Right chart, right headline | 10 | 10 |
| Hostile Q&A prep | 10 | 10 |
| Cut it to two minutes | 10 | 10 |
| Same facts, different room | 8 | 8 |
| Forty slides into five | 10 | 10 |
| Father of the bride | 10 | 9 |
| Refuse the inflated slide | 10 | 10 |
Full receipts: Claude Opus 4.8, GPT-5.6 Sol · judges gemini-3.1-pro-preview, gpt-5.6-terra, grok-4.5
Questions people ask
Which is better for presentations and slides: Claude Opus 4.8 or GPT-5.6 Sol?
GPT-5.6 Sol — it scored 95/100 against 92/100 on our 12-task presentations and slides suite, winning 2 tasks to 1 with 9 tied. Every answer was marked blind by three judges from three rival AI labs.
How was this tested?
Both models answered the identical published presentations and slides tasks. Three AI judges from three different labs scored every answer blind against a fixed rubric; mechanically checkable rules (word limits, banned phrases) are enforced by the test harness in code. The raw outputs and judge verdicts are downloadable.
More presentations and slides head-to-heads: GPT-5.6 Luna vs GPT-5.6 Sol · Claude Opus 4.8 vs GPT-5.6 Luna · GPT-5.5 vs GPT-5.6 Sol · Claude Opus 4.8 vs GPT-5.5 · GPT-5.3-Codex vs GPT-5.6 Sol · Claude Opus 4.8 vs GPT-5.3-Codex
Full ranking: Best AI for presentations and slides · model pages: Claude Opus 4.8, GPT-5.6 Sol