Organization score is the sum of its models
OpenAI
Benefit
+10.66
Harm
−0.21
Net Good
10.45
Timeline
Models
| Rank | Model | Organization | Benefit | Harm | Net | Events | Last impact |
|---|---|---|---|---|---|---|---|
| 1 | ChatGPT | OpenAI | +2.68 | −0.21 | 2.47 | 2 | 22 Jun 2023 |
| 2 | OpenAI Codex | OpenAI | +2.45 | 0.00 | 2.45 | 1 | 21 Jun 2022 |
| 3 | OpenAI Internal Research System — 2026-09 GPT-6 · internal / unreleased | OpenAI | +2.00 | 0.00 | 2.00 | 2 | 8 Sep 2026 |
| 4 | codex-1 | OpenAI | +1.11 | 0.00 | 1.11 | 1 | 16 May 2025 |
| 5 | OpenAI Internal Research Model — 2025-07 OpenAI reasoning models · internal / unreleased | OpenAI | +0.74 | 0.00 | 0.74 | 1 | 21 Jul 2025 |
| 6 | GPT-5.6 Sol | OpenAI | +0.68 | 0.00 | 0.68 | 1 | 8 Sep 2026 |
| 7 | GPT-6 Astra | OpenAI | +0.60 | 0.00 | 0.60 | 3 | 8 Sep 2026 |
| 8 | GPT-4 | OpenAI | +0.40 | 0.00 | 0.40 | 2 | 8 May 2024 |
Domains
Events
OpenAI claims a forced Navier–Stokes singularity from an unreleased system
On 8 September 2026 OpenAI published a 166-page writeup and Lean formalization arguing that smooth, finite-energy 3D Navier–Stokes dynamics with a smooth external force can blow up in finite time, which it says settles Clay statements C and D. The search model was an unreleased internal system; GPT-6 Astra did the Lean step.
Buckmaster and Alpöge, using Sol and Claude, prove smooth-forced blowup for Euler, IPM, and Boussinesq
On 8–11 September 2026 Tristan Buckmaster and Levent Alpöge released Lean-backed finite-time blowup results with smooth forcing for incompressible porous media, 2D Boussinesq, and 3D Euler, crediting Anthropic’s Claude and especially OpenAI’s GPT-5.6 Sol. Astra was used only for writeups and auditing.
An unreleased OpenAI system produces an unforced 3D Euler blowup
While probing Millennium problems in early September 2026, OpenAI says a multi-agent run of an internal model more capable than GPT-6 Astra produced a finite-time singularity for unforced 3D incompressible Euler in about 50 hours.
GPT-6 Astra improves short and large prime-gap bounds
On 3 September 2026 OpenAI released GPT-6 Astra with two analytic number-theory results: a proof that infinitely many consecutive primes differ by at most 186, and an improvement to a large-gap bound that had been stuck for more than 80 years.
Gemini Deep Think and an OpenAI reasoning model both reach IMO 2025 gold-medal scores
In July 2025, Google DeepMind’s Gemini Deep Think was officially graded at 35/42 on the IMO 2025 paper, and OpenAI reported the same score from an experimental reasoning model graded by former medalists.
Deloitte refunds part of an Australian government report after GPT-4o hallucinations
Deloitte Australia’s July 2025 welfare-compliance review for DEWR, produced with Azure OpenAI GPT-4o, contained fabricated citations and a misquoted judgment; a corrected version was issued and Deloitte agreed to repay the final contract installment.