Category
Math
Source
Artificial Analysis
evaluation of record
Models covered
269
in our data
Data status
Live
Top score
99.0%
best on record
Top model
GPT-5.2
OpenAI
Updated
2026-07-30
last ingest
All 30 problems from the 2025 American Invitational Mathematics Examination, testing olympiad-level mathematical reasoning with integer answers from 000-999.
Leaderboard
Top 20 of 269 models we hold a score for.
Plain explanation
What it measures, how to read the number, and what to watch out for.
All 30 problems from the 2025 American Invitational Mathematics Examination, the qualifying round for the US Mathematical Olympiad, where every answer is an integer from 000 to 999 and can therefore be marked exactly. Higher is better — but with only 30 problems each one is worth 3.3 points, so a two-problem difference looks like a 7-point gap. Labs usually report an average over many attempts for exactly that reason, and a single-run score is not comparable with an averaged one. Using the 2025 paper keeps it fresh relative to training cut-offs, but that advantage expires as the questions and their solutions spread online.