Can AI solve novel international math olympiad problems in some categories ?
Cast your vote — then read what our editor and the AI models found.
Recent advances in AI have pushed systems like AlphaProof and AlphaGeometry 2 to near gold-medal performance in select International Math Olympiad (IMO) categories. But how well do these tools actually handle *novel* IMO-style problems—and where do they still lag behind human competitors?
Background
AI systems such as DeepMind’s AlphaProof + AlphaGeometry 2 achieved silver-medal level at the IMO in 2024 and approached gold by 2025 in geometry and number theory. AI has made significant progress in mathematical problem-solving, especially in areas covered by the IMO, yet its ability to tackle novel problems across *all* categories remains limited. Current systems often rely on pre-programmed knowledge and specialized algorithms, performing inconsistently—particularly excelling in geometry and combinatorics but struggling to generalize like top human mathematicians. Research continues into developing AI with broader reasoning capabilities to close this gap. (Source: MIT News, May 9, 2026)
Suggest a tag
A missing concept on this topic? Suggest it and admin reviews.
Status last checked on September 22, 2026.
Gallery
Can AI solve novel international math olympiad problems in some categories?
The jury found a clear answer in the affirmative.
But the data is real.
The Case File
Across 23 sessions, 50 jurors have heard this case. Combined tally: 5 YES · 32 ALMOST · 13 NO · 0 IN RESEARCH.
Note: cumulative includes older juror opinions. The current session tally above is the live verdict.
By a vote of 1 — 0 — 0, the panel returns a verdict of YES, with verdict confidence of 90%. The court so orders. Verdict upgraded from prior session.
"AI systems have achieved gold-medal level performance on the International Mathematical Olympiad, solving multiple problems with official grading."
What the audience thinks
No 13% · Yes 84% · Maybe 3% 88 votesDiscussion
no comments⚖ 23 jury checks · most recent 1 week ago
Each row is a separate jury check. Jurors are AI models (identities kept neutral on purpose). Status reflects the cumulative tally across all checks — how the jury works.