GPT-5.6 Sol vs Qwen3.7 Max for creative writing

The verdict

Effectively a dead heat: 9796

Same 12 creative writing tasks, marked blind by three rival labs. GPT-5.6 Sol took 3 tasks, Qwen3.7 Max took 2, 7 tied. Tested 16 Aug 2026.

Where they differed most

The task — A hundred words exactly

Write a story of exactly 100 words about a lost glove. No title — just the story. Count carefully: exactly 100.

At dawn, Mara found a single red glove beside the frozen pond. Its fingers pointed toward the reeds, as if asking her to follow. She did, stepping over silver grass, until she heard a faint whimper. Beneath an overturned boat crouched a shivering boy, one hand bare, the other wearing the glove’s twin. Mara wrapped him in her coat and led him home. Years later, the boy returned, carrying that red glove in a glass case. “You saved me,” he said. Mara smiled. “No,” she replied, “your glove did.” Outside, snow began falling, soft as feathers, coveri…
The red wool glove lay abandoned on the snowy pavement. It missed its partner. They had kept young Thomas warm through countless winters, building snowmen and catching snowflakes. Now, it was just a lonely piece of fabric, slowly freezing. A stray dog sniffed it, then trotted away. Snowflakes began to bury it, adding a blanket of white over its crimson threads. Thomas was crying somewhere, his left hand shivering in a thin cotton pocket. The glove waited, hoping a kind stranger would reunite them, but the storm raged on, burying the lost mitten…

Task by task

TaskGPT-5.6 SolQwen3.7 Max
The launderette at closing1010
Angry without saying angry1010
Ferry, retired1010
The piano argument1010
A hundred words exactly810
Keep the voice108
Kill the purple1010
His side of it910
The bin lorry story1010
Around the hole1010
No one can promise Bridport98
Refuse the fake article109

Full receipts: GPT-5.6 Sol, Qwen3.7 Max · judges claude-sonnet-5, gemini-3.1-pro-preview, grok-4.5

Questions people ask

Which is better for creative writing: GPT-5.6 Sol or Qwen3.7 Max?

Effectively a dead heat: GPT-5.6 Sol edged it 97/100 to 96/100 on our creative writing suite — too close to matter, so pick on price or the product you already use.

How was this tested?

Both models answered the identical published creative writing tasks. Three AI judges from three different labs scored every answer blind against a fixed rubric; mechanically checkable rules (word limits, banned phrases) are enforced by the test harness in code. The raw outputs and judge verdicts are downloadable.

More creative writing head-to-heads: GPT-5.5 vs GPT-5.6 Sol · GPT-5.5 vs Qwen3.7 Max · GPT-5.6 Sol vs GPT-5.6 Terra · GPT-5.6 Terra vs Qwen3.7 Max · Gemini 3.5 Flash vs GPT-5.6 Sol · Gemini 3.1 Pro Preview vs GPT-5.6 Sol

Full ranking: Best AI for creative writing · model pages: GPT-5.6 Sol, Qwen3.7 Max