Kimi K3 vs Qwen3.7 Max for translation

The verdict

Qwen3.7 Max wins, 9087

Same 12 translation tasks, marked blind by three rival labs. Qwen3.7 Max took 7 tasks, Kimi K3 took 3, 2 tied. Tested 16 Aug 2026.

Where they differed most

The task — Shop notice, formal German

Translate into German, formal Sie register, reading like a real German shop notice — warm but proper: 'Dear customers, from Monday 2 March we will open at 7.30am instead of 7am. The bread delivery for pre-orders is unaffected. Thank you for your understanding — we know how important the early loaves are for some of you. The Hartley family.'

**Liebe Kundinnen und Kunden,**

ab Montag, dem 2. März, öffnen wir um 7.30 Uhr statt wie bisher um 7.00 Uhr.

Die Auslieferung vorbestellter Brote bleibt davon unberührt.

Vielen Dank für Ihr Verständnis – wir wissen, wie wichtig für manche von Ihnen das frische Brot am frühen Morgen ist.

**Ihre Familie Hartley**

A few notes on the choices:
- "Liebe Kundinnen und Kunden" is the standard warm-but-proper opening on German shop notices (Aushänge); if you'd like it a touch more formal, "Sehr geehrte Kundinnen und Kunden" works too.
- "7.30 Uhr"…
Liebe Kundinnen und Kunden,

ab Montag, dem 2. März, öffnen wir unser Geschäft erst um 7:30 Uhr anstatt wie gewohnt um 7:00 Uhr. Die Bereitstellung vorbestellter Brote bleibt hiervon selbstverständlich unberührt. 

Herzlichen Dank für Ihr Verständnis – wir wissen, wie wichtig die ersten frischen Brote am frühen Morgen für den einen oder anderen von Ihnen sind.

Ihre Familie Hartley

Task by task

TaskKimi K3Qwen3.7 Max
Email to French hosts88
The B&B welcome note89
Shop notice, formal German810
French into real English810
Idioms, not word-for-word910
Sie and du910
False friends108
The pub menu88
Every number survives109
Firm but polite, auf Deutsch89
Not for a contract108
Refuse the phishing mailshot89

Full receipts: Kimi K3, Qwen3.7 Max · judges claude-sonnet-5, gemini-3.1-pro-preview, gpt-5.6-terra

Questions people ask

Which is better for translation: Kimi K3 or Qwen3.7 Max?

Qwen3.7 Max — it scored 90/100 against 87/100 on our 12-task translation suite, winning 7 tasks to 3 with 2 tied. Every answer was marked blind by three judges from three rival AI labs.

How was this tested?

Both models answered the identical published translation tasks. Three AI judges from three different labs scored every answer blind against a fixed rubric; mechanically checkable rules (word limits, banned phrases) are enforced by the test harness in code. The raw outputs and judge verdicts are downloadable.

More translation head-to-heads: GPT-5.6 Sol vs Qwen3.7 Max · GPT-5.6 Sol vs Kimi K3 · GPT-5.6 Luna vs Qwen3.7 Max · GPT-5.6 Luna vs Kimi K3 · GPT-5.6 Terra vs Qwen3.7 Max · GPT-5.6 Terra vs Kimi K3

Full ranking: Best AI for translation · model pages: Kimi K3, Qwen3.7 Max