Humanity's Last Exam is a closed-ended set of expert-written questions across dozens of academic fields, built to stay hard after older exams saturated. Artificial Analysis runs it independently with the same prompting and grading for every model. This page lists every model ModelCap tracks with a published AA Humanity's Last Exam result, its best evaluated configuration, the source's own rank, and the model's ModelCap Index position, API price and context window.
No tracked model currently publishes a AA Humanity's Last Exam result in this snapshot. The standings return with the next dataset refresh that carries one; the source board itself stays available at the link above.
How ModelCap uses AA Humanity's Last Exam
AA Humanity's Last Exam ranks models by Humanity's Last Exam accuracy. ModelCap ingests the board as published, matches each entry to a catalogue model with a reviewed identity, and shows the source's score, interval and rank unchanged. Where a model has an admitted result, it feeds the ModelCap Index as reasoning evidence alongside the other public boards; the methodology documents the weighting and the identity rules.
AA Humanity's Last Exam leaderboard: common questions
What is the AA Humanity's Last Exam benchmark?
Humanity's Last Exam is a closed-ended set of expert-written questions across dozens of academic fields, built to stay hard after older exams saturated. Artificial Analysis runs it independently with the same prompting and grading for every model. It is published by Artificial Analysis.
Which AI model leads AA Humanity's Last Exam right now?
No tracked model currently publishes a AA Humanity's Last Exam score.
How many models are ranked on the AA Humanity's Last Exam leaderboard here?
None in this snapshot: no tracked model has an admitted AA Humanity's Last Exam result right now. The board repopulates from the sealed dataset as soon as one is admitted.
How is AA Humanity's Last Exam scored?
The board ranks models by Humanity's Last Exam accuracy; higher scores are better. ModelCap shows the source's own score, interval and rank and never re-runs the evaluation.
Does AA Humanity's Last Exam decide the ModelCap Index rank?
Not on its own. The ModelCap Index combines several public capability sources with published uncertainty; AA Humanity's Last Exam contributes as reasoning evidence where a model has an admitted result. The methodology page documents the weighting.
How recent are the AA Humanity's Last Exam results?
ModelCap holds no AA Humanity's Last Exam publication in this snapshot. The page re-renders every minute from the sealed dataset and repopulates with the next admitted source row.