IFBench

A bird's-eye view of the benchmark: what it measures and every AI IQ chart built on it.

About IFBench

IFBench has source-backed multi-effort configurations. The Cost Efficiency chart expands only those configurations; the score bar remains one canonical row per model.

IFBench Cost Efficiency
Source-matched reasoning-effort score and task-cost configurations. Each line is one canonical model. Color = provider.

How to read this chart

Each line uses the shared model-configuration dataset to compare source-matched reasoning-effort levels on IFBench. Every point shows the score and AA cost per task for that exact configuration. Canonical score bars and bell curves remain one point per model.

IFBench Benchmark Scores
Instruction-following and constraint-adherence scores. Color = provider.

How to read this chart

Bars rank models by the source-backed benchmark value used for this chart. Longer bars indicate higher published scores.

IFBench vs Effective Cost
X = effective cost (log). Y = IFBench %. Color = provider.
Controls:
IFBench1:1Cost

How to read this chart

Each point is a public model. The chart compares IFBench % against Effective Cost (per 1M I/O Tokens), with color showing the model provider.