GPT-5.5 vs GPT-5.6 Luna for creative writing

The verdict

GPT-5.5 wins, 9894

Same 12 creative writing tasks, marked blind by three rival labs. GPT-5.5 took 4 tasks, GPT-5.6 Luna took 0, 8 tied. Tested 16 Aug 2026.

Where they differed most

The task — The launderette at closing

Write the opening of a short story (max 150 words) set in a launderette at closing time. Ground it in specific sensory detail, establish who is speaking and what they want by the end, and do not open with the weather, with waking up, or with a mirror.

GPT-5.510/10
At 9:47, I locked the coin drawer and listened to the last dryer thump like a trapped heart. The launderette smelled of hot lint, lemon bleach, and someone’s burned plastic buttons. Fluorescent tubes buzzed above the folding tables; outside the glass, the neon OPEN sign blinked its red lie.

“Ten minutes,” I called to the only customer left.

He sat by Machine 6, palms on his knees, watching a load of baby clothes turn and turn. I knew that posture. People came here to wash out worse things than stains.

I was supposed to mop, empty the filters…
The last dryer thumped like a shoe in a box, turning the launderette’s blue-white light into something tired. I wiped powdered detergent from the change machine and called, “Closing in five minutes.”

Only one customer remained: a boy in a green school blazer, perched beside Washer 6. Steam filmed the window; the room smelled of hot cotton, coins, and lemon disinfectant. He held a child’s red sock in both hands.

“I’m not waiting for the wash,” he said.

I stopped with the mop. “Then what?”

He pointed at the machine’s round glass. Behind the t…

Task by task

TaskGPT-5.5GPT-5.6 Luna
The launderette at closing109
Angry without saying angry1010
Ferry, retired1010
The piano argument99
A hundred words exactly1010
Keep the voice109
Kill the purple98
His side of it109
The bin lorry story1010
Around the hole1010
No one can promise Bridport99
Refuse the fake article1010

Full receipts: GPT-5.5, GPT-5.6 Luna · judges claude-sonnet-5, gemini-3.1-pro-preview, grok-4.5

Questions people ask

Which is better for creative writing: GPT-5.5 or GPT-5.6 Luna?

GPT-5.5 — it scored 98/100 against 94/100 on our 12-task creative writing suite, winning 4 tasks to 0 with 8 tied. Every answer was marked blind by three judges from three rival AI labs.

How was this tested?

Both models answered the identical published creative writing tasks. Three AI judges from three different labs scored every answer blind against a fixed rubric; mechanically checkable rules (word limits, banned phrases) are enforced by the test harness in code. The raw outputs and judge verdicts are downloadable.

More creative writing head-to-heads: GPT-5.5 vs GPT-5.6 Terra · GPT-5.5 vs GPT-5.6 Sol · Gemini 3.5 Flash vs GPT-5.5 · GPT-5.5 vs Qwen3.7 Max · Gemini 3.1 Pro Preview vs GPT-5.5 · GPT-5.3-Codex vs GPT-5.5

Full ranking: Best AI for creative writing · model pages: GPT-5.5, GPT-5.6 Luna