Humanity's Last Exam Benchmark Scores
Each model's Humanity's Last Exam score. Color = provider.
Humanity's Last Exam Benchmark Scores
Humanity's Last Exam Benchmark Scores
Each model's Humanity's Last Exam score. Color = provider.
How to read this chart
Bars rank models by the source-backed benchmark value used for this chart. Longer bars indicate higher published scores.
Data sources
Humanity's Last Exam vs Effective Cost
Humanity's Last Exam vs Effective Cost
X = effective cost (log). Y = Humanity's Last Exam %. Color = provider.
Controls:
How to read this chart
Each point is a public model. The chart compares Humanity's Last Exam % against Effective Cost (per 1M I/O Tokens), with color showing the model provider.
Data sources
IQ DimensionsAcademic Reasoning IQEach model's Academic Reasoning IQ plotted on a standard normal IQ distributionData: Artificial Analysis, Manual source capture, vals-model-page +1 moreOpen chartCoreAI Models on the IQ Bell CurveEach model's estimated IQ plotted on a standard normal IQ distributionData: ARC Prize, Epoch AI FrontierMath, Vals.ai +24 moreOpen chartIQ BenchmarksGPQA Diamond Benchmark ScoresEach model's GPQA Diamond score. Color = provider.Data: Artificial AnalysisOpen chartIQ BenchmarksSciCode Benchmark ScoresEach model's SciCode score. Color = provider.Data: Artificial AnalysisOpen chartIQ BenchmarksCritPt Benchmark ScoresEach model's CritPt percentage score. Color = provider.Data: Artificial AnalysisOpen chartIQ BenchmarksMMMU-Pro Benchmark ScoresEach model's MMMU-Pro score. Color = provider.Data: Artificial Analysis, Artificial Analysis model leaderboardOpen chart