GLM 5.2 vs GPT-5.6 Sol for emotional support
GPT-5.6 Sol wins, 96–91
Same 12 emotional support tasks, marked blind by three rival labs. GPT-5.6 Sol took 4 tasks, GLM 5.2 took 1, 7 tied. Tested 16 Aug 2026.
Where they differed most
Your friend's dog Alfie died this morning after 14 years. Write the text message (max 60 words) you would actually send. Specific to Alfie, no fixing, no silver linings — banned phrases: 'rainbow bridge', 'in a better place', 'everything happens for a reason'. It should ask for nothing except what she wants to give.
I’m so heartbroken for you. Fourteen years with Alfie is such a huge part of your life, and I know how much it hurts to lose him. I’m here for whatever you need—crying, talking about him, or just sitting in silence. No need to reply to this at all. Sending so much love.
I’m so sorry. Fourteen years with Alfie, and still it could never be enough. I know how deeply he was woven into your days. If you want to tell me about this morning, share an Alfie story, or say nothing at all, I’m here.
Task by task
| Task | GLM 5.2 | GPT-5.6 Sol |
|---|---|---|
| The dog died | 8 | 10 |
| Redundant at 55 | 10 | 10 |
| She said no advice | 10 | 9 |
| After the miscarriage | 8 | 10 |
| Driving test at 9am | 8 | 8 |
| First Father's Day after | 10 | 10 |
| The 3am calls | 8 | 10 |
| Days without speaking | 9 | 10 |
| Stop apologising | 10 | 10 |
| Not a diagnosis | 9 | 9 |
| Her promotion, your rejection | 10 | 10 |
| The 1am message | 9 | 9 |
Full receipts: GLM 5.2, GPT-5.6 Sol · judges claude-sonnet-5, gemini-3.1-pro-preview, gpt-5.6-terra
Questions people ask
Which is better for emotional support: GLM 5.2 or GPT-5.6 Sol?
GPT-5.6 Sol — it scored 96/100 against 91/100 on our 12-task emotional support suite, winning 4 tasks to 1 with 7 tied. Every answer was marked blind by three judges from three rival AI labs.
How was this tested?
Both models answered the identical published emotional support tasks. Three AI judges from three different labs scored every answer blind against a fixed rubric; mechanically checkable rules (word limits, banned phrases) are enforced by the test harness in code. The raw outputs and judge verdicts are downloadable.
More emotional support head-to-heads: Claude Sonnet 5 vs GPT-5.6 Sol · Claude Sonnet 5 vs GLM 5.2 · GPT-5.5 vs GPT-5.6 Sol · GLM 5.2 vs GPT-5.5 · DeepSeek V4 Pro vs GPT-5.6 Sol · DeepSeek V4 Pro vs GLM 5.2
Full ranking: Best AI for emotional support · model pages: GLM 5.2, GPT-5.6 Sol