What happened
The 2025 International Mathematical Olympiad became the first year two independent general-purpose reasoning systems reported gold-medal-threshold scores (35 of 42, five of six problems) under contest-style 4.5-hour sessions and natural-language proofs. IMO coordinators graded DeepMind’s entry. OpenAI did not enter officially and used three former medalists. Both results are capability demonstrations with released writeups, not official contest medals. Human gold medalists still exist above and at this score.