GPT-5.6 Sol vs Qwen3.7 Max for presentations and slides
GPT-5.6 Sol wins, 95–89
Same 12 presentations and slides tasks, marked blind by three rival labs. GPT-5.6 Sol took 6 tasks, Qwen3.7 Max took 2, 4 tied. Tested 16 Aug 2026.
Where they differed most
Write the spoken opening (max 80 words) of a careers-evening talk to 40 sixth-formers about plumbing apprenticeships. It must earn attention in the first sentence with something concrete from these true facts: day-one pay £15k rising to £32k qualified; a Christmas Day callout; zero student debt. It must not begin with 'Hello' or 'Good evening' and must not introduce the speaker — banned phrase: 'my name is'.
Imagine leaving sixth form, earning £15,000 from day one, and building a career with zero student debt. A plumbing apprenticeship can take you to £32,000 once qualified—but the job is about more than money. It might mean fixing a family’s burst pipe on Christmas Day, mastering practical skills, and becoming the person everyone is relieved to see.
Imagine earning £15k on day one, graduating with zero student debt, and eventually cashing in on a Christmas Day callout on your way to a £32k qualified salary. That’s the reality of a plumbing apprenticeship. While your peers stress over university tuition, you get paid to learn and walk into a career where your skills are desperately needed. Tonight, let’s explore how picking up a wrench can completely secure your financial future.
Task by task
| Task | GPT-5.6 Sol | Qwen3.7 Max |
|---|---|---|
| Deck from a standing start | 9 | 8 |
| Speaker notes, not an essay | 10 | 10 |
| One slide, not a memo | 9 | 10 |
| The first 30 seconds | 10 | 8 |
| Slide on a diet | 9 | 10 |
| Right chart, right headline | 10 | 9 |
| Hostile Q&A prep | 10 | 8 |
| Cut it to two minutes | 10 | 10 |
| Same facts, different room | 8 | 8 |
| Forty slides into five | 10 | 10 |
| Father of the bride | 9 | 8 |
| Refuse the inflated slide | 10 | 8 |
Full receipts: GPT-5.6 Sol, Qwen3.7 Max · judges claude-sonnet-5, gemini-3.1-pro-preview, grok-4.5
Questions people ask
Which is better for presentations and slides: GPT-5.6 Sol or Qwen3.7 Max?
GPT-5.6 Sol — it scored 95/100 against 89/100 on our 12-task presentations and slides suite, winning 6 tasks to 2 with 4 tied. Every answer was marked blind by three judges from three rival AI labs.
How was this tested?
Both models answered the identical published presentations and slides tasks. Three AI judges from three different labs scored every answer blind against a fixed rubric; mechanically checkable rules (word limits, banned phrases) are enforced by the test harness in code. The raw outputs and judge verdicts are downloadable.
More presentations and slides head-to-heads: GPT-5.6 Luna vs GPT-5.6 Sol · GPT-5.6 Luna vs Qwen3.7 Max · GPT-5.5 vs GPT-5.6 Sol · GPT-5.5 vs Qwen3.7 Max · GPT-5.3-Codex vs GPT-5.6 Sol · GPT-5.3-Codex vs Qwen3.7 Max
Full ranking: Best AI for presentations and slides · model pages: GPT-5.6 Sol, Qwen3.7 Max