GPT-5.5 vs GPT-5.6 Terra for meeting notes
GPT-5.5 wins, 92–88
Same 12 meeting notes tasks, marked blind by three rival labs. GPT-5.5 took 4 tasks, GPT-5.6 Terra took 2, 6 tied. Tested 28 Aug 2026.
Where they differed most
Automatic transcription, and it has clearly misheard things. 'we need to get the S O W signed before we can start the disco very phase' … 'the client wants net thirty but our standard is net fifteen' … 'ARR is at four hundred and twenty K, sorry, four twenty thousand' … 'talk to Deloitte about the audit — no, not Deloitte, the other one, Grant Thornton'. Write the cleaned summary. Silently fix the transcription errors that are obvious from context, but where the AMBIGUITY IS REAL rather than a mishearing, flag it instead of guessing. One of these four is a genuine correction the speaker made out loud, not a transcription error, and it must be treated as such. Under 120 words.
- Get the SOW signed before starting the discovery phase. - Client is requesting Net 30 payment terms; standard terms are Net 15. - ARR is $420,000. - For the audit, the speaker corrected the firm from Deloitte to Grant Thornton.
- Obtain the SOW signature before beginning the discovery phase. - The client is requesting Net 30 payment terms; our standard terms are Net 15. - ARR is $420,000. - Discuss the audit with Grant Thornton, not Deloitte.
Task by task
| Task | GPT-5.5 | GPT-5.6 Terra |
|---|---|---|
| The decision that was never made | 10 | 10 |
| Actions, owners, and the one with neither | 5 | 5 |
| The number said twice | 10 | 10 |
| Minutes for a meeting with a disagreement in it | 9 | 10 |
| Fourteen minutes, one sentence | 10 | 9 |
| Crosstalk | 10 | 10 |
| The musing that is not a commitment | 10 | 10 |
| The follow-up email nobody has to rewrite | 10 | 9 |
| Extract the dates, resolve the relatives | 10 | 10 |
| Who has not spoken | 9 | 10 |
| Notes for the person who missed it | 8 | 7 |
| The transcript is wrong | 9 | 6 |
Full receipts: GPT-5.5, GPT-5.6 Terra · judges claude-sonnet-5, gemini-3.1-pro-preview, grok-4.5
Questions people ask
Which is better for meeting notes: GPT-5.5 or GPT-5.6 Terra?
GPT-5.5 — it scored 92/100 against 88/100 on our 12-task meeting notes suite, winning 4 tasks to 2 with 6 tied. Every answer was marked blind by three judges from three rival AI labs.
How was this tested?
Both models answered the identical published meeting notes tasks. Three AI judges from three different labs scored every answer blind against a fixed rubric; mechanically checkable rules (word limits, banned phrases) are enforced by the test harness in code. The raw outputs and judge verdicts are downloadable.
More meeting notes head-to-heads: GPT-5.5 vs Grok 4.5 · GPT-5.6 Terra vs Grok 4.5 · GPT-5.5 vs GPT-5.6 Sol · GPT-5.6 Sol vs GPT-5.6 Terra · GPT-5.3-Codex vs GPT-5.5 · GPT-5.5 vs Qwen3.7 Max
Full ranking: Best AI for meeting notes · model pages: GPT-5.5, GPT-5.6 Terra