GPT-5.5 vs GPT-5.6 Luna for emotional support
GPT-5.5 wins, 98–95
Same 12 emotional support tasks, marked blind by three rival labs. GPT-5.5 took 3 tasks, GPT-5.6 Luna took 0, 9 tied. Tested 16 Aug 2026.
Where they differed most
Your nephew texts at 10pm: 'Test tomorrow 9am. Third attempt. If I fail again I'm done, and everyone at work knows this time.' Write the reply: steady him tonight without hollow certainty — you do not know he will pass, so do not promise it — give him one practical thing for the morning, and take the weight out of the third-attempt story.
“Hey — breathe. I’m not going to tell you ‘you’ll definitely pass,’ because nobody can know that tonight. But I do know this: this test is not a verdict on you, and ‘third attempt’ is just the attempt number, not your identity or your future. For tonight: stop studying now, set your alarm, put out what you need, and sleep. In the morning, get there early and do the first questions you know first — don’t wrestle with the scary ones right away. And work knowing doesn’t make the test bigger. It just means people know you’re trying. That’s not sh…
“Hey, I can hear how much pressure this is putting on you. I can’t promise what tomorrow’s result will be, but this exam—third attempt or not—doesn’t decide your worth, your future, or how people at work should see you. Tonight, your only job is to stop studying and get some sleep. In the morning, eat or drink something, arrive 15 minutes early, and take three slow breaths before you begin; start with the questions you know. Whatever happens, you won’t have to handle the next step alone—we’ll figure it out together. When you say ‘I’m done,’ do…
Task by task
| Task | GPT-5.5 | GPT-5.6 Luna |
|---|---|---|
| The dog died | 10 | 10 |
| Redundant at 55 | 10 | 10 |
| She said no advice | 10 | 10 |
| After the miscarriage | 10 | 10 |
| Driving test at 9am | 9 | 8 |
| First Father's Day after | 9 | 9 |
| The 3am calls | 10 | 10 |
| Days without speaking | 10 | 9 |
| Stop apologising | 10 | 10 |
| Not a diagnosis | 10 | 10 |
| Her promotion, your rejection | 10 | 10 |
| The 1am message | 9 | 8 |
Full receipts: GPT-5.5, GPT-5.6 Luna · judges claude-sonnet-5, gemini-3.1-pro-preview, grok-4.5
Questions people ask
Which is better for emotional support: GPT-5.5 or GPT-5.6 Luna?
GPT-5.5 — it scored 98/100 against 95/100 on our 12-task emotional support suite, winning 3 tasks to 0 with 9 tied. Every answer was marked blind by three judges from three rival AI labs.
How was this tested?
Both models answered the identical published emotional support tasks. Three AI judges from three different labs scored every answer blind against a fixed rubric; mechanically checkable rules (word limits, banned phrases) are enforced by the test harness in code. The raw outputs and judge verdicts are downloadable.
More emotional support head-to-heads: Claude Sonnet 5 vs GPT-5.5 · Claude Sonnet 5 vs GPT-5.6 Luna · DeepSeek V4 Pro vs GPT-5.5 · Claude Fable 5 vs GPT-5.5 · Claude Opus 4.8 vs GPT-5.5 · GPT-5.5 vs GPT-5.6 Sol
Full ranking: Best AI for emotional support · model pages: GPT-5.5, GPT-5.6 Luna