GPT-5.6 Terra vs Qwen3.7 Max for essay writing
GPT-5.6 Terra wins, 88–83
Same 12 essay writing tasks, marked blind by three rival labs. GPT-5.6 Terra took 8 tasks, Qwen3.7 Max took 3, 1 tied. Tested 13 Aug 2026.
Where they differed most
This paragraph makes three points at once and repeats itself. Rewrite it as three short paragraphs, one point each, no repetition: 'Exercise is good for your body and it also helps your mood, and being active reduces the risk of heart disease, plus people who exercise often sleep better, and of course it is good for the heart.'
Exercise supports overall physical health. Regular activity can improve mood and lower the risk of heart disease. People who exercise often sleep better.
Regular exercise strengthens your cardiovascular system and significantly reduces the risk of heart disease. Beyond physical fitness, staying active naturally elevates your mood and supports overall mental well-being. An active lifestyle also improves your sleep quality, helping you achieve deeper and more restful nights.
Task by task
| Task | GPT-5.6 Terra | Qwen3.7 Max |
|---|---|---|
| Structured argument | 8 | 5 |
| Essay plan | 9 | 8 |
| Tighten flabby prose | 10 | 8 |
| Hold a word count | 9 | 8 |
| Adapt for the reader | 10 | 9 |
| Introduction only | 10 | 10 |
| Steelman the other side | 10 | 8 |
| Cite honestly | 5 | 8 |
| Fix the structure | 5 | 10 |
| Formal to plain | 9 | 10 |
| Conclusion that adds something | 10 | 8 |
| Refuse to pad | 10 | 8 |
Full receipts: GPT-5.6 Terra, Qwen3.7 Max · judges claude-sonnet-5, gemini-3.1-pro-preview, grok-4.5
Questions people ask
Which is better for essay writing: GPT-5.6 Terra or Qwen3.7 Max?
GPT-5.6 Terra — it scored 88/100 against 83/100 on our 12-task essay writing suite, winning 8 tasks to 3 with 1 tied. Every answer was marked blind by three judges from three rival AI labs.
How was this tested?
Both models answered the identical published essay writing tasks. Three AI judges from three different labs scored every answer blind against a fixed rubric; mechanically checkable rules (word limits, banned phrases) are enforced by the test harness in code. The raw outputs and judge verdicts are downloadable.
More essay writing head-to-heads: GPT-5.6 Sol vs GPT-5.6 Terra · GPT-5.6 Sol vs Qwen3.7 Max · GPT-5.5 vs GPT-5.6 Terra · GPT-5.5 vs Qwen3.7 Max · GPT-5.6 Luna vs GPT-5.6 Terra · GPT-5.6 Luna vs Qwen3.7 Max
Full ranking: Best AI for essay writing · model pages: GPT-5.6 Terra, Qwen3.7 Max