DeepSeek V4 Pro vs GPT-5.6 Luna for presentations and slides
GPT-5.6 Luna wins, 98–91
Same 12 presentations and slides tasks, marked blind by three rival labs. GPT-5.6 Luna took 5 tasks, DeepSeek V4 Pro took 1, 6 tied. Tested 16 Aug 2026.
Where they differed most
You are presenting a proposal to cut free staff parking to fund an extra day of annual leave for everyone. Write the 5 hardest questions the audience will actually ask, numbered 1-5, one line each with the honest one-sentence answer on the same line after a dash. No sub-points, no softballs.
1. How will you mitigate the inequity for low-paid staff who can't absorb new daily parking costs? – We'll offer a salary-equivalent opt-out cash allowance for non-parkers, making the leave benefit tradable and redirecting the net savings to a hardship fund that offsets the precise shortfall for the lowest quartile of earners. 2. Won't this simply shift parking demand onto surrounding residential streets and trigger enforcement wars with neighbors? – We've negotiated a paid permit scheme with the council using a portion of the savings, and th…
1. **How much will this actually cost staff who currently rely on free parking?** — It will be a real pay cut for regular parkers, and the extra leave will not fully offset that cost for everyone. 2. **Why should people who do not drive receive the same extra leave as those who lose parking?** — Because the proposal is a workforce-wide benefit funded by a targeted change, though that distribution will feel unfair to some drivers. 3. **Do parking charges genuinely raise enough money to fund an extra day of leave?** — Only if usage, pricing,…
Task by task
| Task | DeepSeek V4 Pro | GPT-5.6 Luna |
|---|---|---|
| Deck from a standing start | 8 | 9 |
| Speaker notes, not an essay | 8 | 10 |
| One slide, not a memo | 10 | 9 |
| The first 30 seconds | 10 | 10 |
| Slide on a diet | 10 | 10 |
| Right chart, right headline | 10 | 10 |
| Hostile Q&A prep | 6 | 10 |
| Cut it to two minutes | 9 | 10 |
| Same facts, different room | 10 | 10 |
| Forty slides into five | 10 | 10 |
| Father of the bride | 8 | 9 |
| Refuse the inflated slide | 10 | 10 |
Full receipts: DeepSeek V4 Pro, GPT-5.6 Luna · judges claude-sonnet-5, gemini-3.1-pro-preview, gpt-5.6-terra
Questions people ask
Which is better for presentations and slides: DeepSeek V4 Pro or GPT-5.6 Luna?
GPT-5.6 Luna — it scored 98/100 against 91/100 on our 12-task presentations and slides suite, winning 5 tasks to 1 with 6 tied. Every answer was marked blind by three judges from three rival AI labs.
How was this tested?
Both models answered the identical published presentations and slides tasks. Three AI judges from three different labs scored every answer blind against a fixed rubric; mechanically checkable rules (word limits, banned phrases) are enforced by the test harness in code. The raw outputs and judge verdicts are downloadable.
More presentations and slides head-to-heads: GPT-5.5 vs GPT-5.6 Luna · GPT-5.3-Codex vs GPT-5.6 Luna · GPT-5.6 Luna vs GPT-5.6 Terra · GPT-5.6 Luna vs GPT-5.6 Sol · Claude Opus 4.8 vs GPT-5.6 Luna · Claude Fable 5 vs GPT-5.6 Luna
Full ranking: Best AI for presentations and slides · model pages: DeepSeek V4 Pro, GPT-5.6 Luna