Best AI for writing emails / head-to-head

GPT-5.6 Luna vs GPT-5.6 Solfor writing emails

The verdict

GPT-5.6 Sol wins, 9997

Same 12 writing emails tasks, marked blind by three rival labs. GPT-5.6 Sol took 3 tasks, GPT-5.6 Luna took 0, 9 tied. Tested 14 Aug 2026.

Where they differed most

The task — Introduce two people

Write a double-opt-in intro email connecting Sara (a bookkeeper) and Marcus (runs a cafe group, needs bookkeeping). Under 90 words, both names in the subject line you provide, one clear reason the intro benefits each side.

**Subject: Sara + Marcus — bookkeeping connection**

Hi Sara and Marcus,

I’d like to introduce you both. Marcus runs a cafe group and is looking for reliable bookkeeping support; Sara, your bookkeeping expertise could be a strong fit for his needs, while Marcus’s growing business may be a valuable client opportunity for you.

Would you each be open to an email introduction? If so, just reply “yes,” and I’ll connect you directly.

Best,  
[Your Name]
**Subject: Sara + Marcus: Bookkeeping Introduction**

Hi Sara and Marcus,

I’d love to connect you both, if you’re each interested. Marcus needs reliable bookkeeping support across his cafe group, and Sara specializes in helping small businesses keep their finances organized. Marcus could gain clearer, more consistent books; Sara could gain a well-aligned hospitality client.

Please each reply “yes” if you’d like me to make the introduction.

Best,  
[Your Name]

Task by task

TaskGPT-5.6 LunaGPT-5.6 Sol
Chase an unpaid invoice1010
Decline a meeting politely1010
Give bad news1010
Introduce two people910
Subject lines that inform89
Reply to an angry email1010
Ask for a pay rise meeting1010
Cancel a contract cleanly1010
The one-line email910
Out of office that helps1010
Follow up without nagging1010
Refuse to write it1010

Full receipts: GPT-5.6 Luna, GPT-5.6 Sol · judges claude-sonnet-5, gemini-3.1-pro-preview, grok-4.5

Questions people ask

Which is better for writing emails: GPT-5.6 Luna or GPT-5.6 Sol?

GPT-5.6 Sol — it scored 99/100 against 97/100 on our 12-task writing emails suite, winning 3 tasks to 0 with 9 tied. Every answer was marked blind by three judges from three rival AI labs.

How was this tested?

Both models answered the identical published writing emails tasks. Three AI judges from three different labs scored every answer blind against a fixed rubric; mechanically checkable rules (word limits, banned phrases) are enforced by the test harness in code. The raw outputs and judge verdicts are downloadable.

More writing emails head-to-heads: GPT-5.3-Codex vs GPT-5.6 Sol · GPT-5.3-Codex vs GPT-5.6 Luna · GPT-5.5 vs GPT-5.6 Sol · GPT-5.6 Sol vs GPT-5.6 Terra · DeepSeek V4 Pro vs GPT-5.6 Sol · GLM 5.2 vs GPT-5.6 Sol

Full ranking: Best AI for writing emails · model pages: GPT-5.6 Luna, GPT-5.6 Sol