CritPt

A bird's-eye view of the benchmark: what it measures and every AI IQ chart built on it.

About CritPt

CritPt has source-backed multi-effort configurations. The Cost Efficiency chart expands only those configurations; the score bar remains one canonical row per model.

CritPt Cost Efficiency
Source-matched reasoning-effort score and task-cost configurations. Each line is one canonical model. Color = provider.

How to read this chart

Each line uses the shared model-configuration dataset to compare source-matched reasoning-effort levels on CritPt. Every point shows the score and AA cost per task for that exact configuration. Canonical score bars and bell curves remain one point per model.

CritPt Benchmark Scores
Each model's CritPt percentage score. Color = provider.

How to read this chart

Bars rank models by the source-backed benchmark value used for this chart. Longer bars indicate higher published scores.

CritPt vs Effective Cost
X = effective cost (log). Y = CritPt %. Color = provider.
Controls:
CritPt1:1Cost

How to read this chart

Each point is a public model. The chart compares CritPt % against Effective Cost (per 1M I/O Tokens), with color showing the model provider.