Animating Cost over IQ vs Completion Time
X = completion time (time to last token, TTLT; log, reversed). Y = IQ. Effective Cost/1M I/O Tokens is shown as bills over a 10s period.
Animating Cost over IQ vs Completion Time
Animating Cost over IQ vs Completion Time
X = completion time (time to last token, TTLT; log, reversed). Y = IQ. Effective Cost/1M I/O Tokens is shown as bills over a 10s period.
How to read this chart
Animated bills show relative spend over time while the axes compare quality, cost, or response speed.
Data sources
Response TimeIQ vs ThroughputThroughput is median output tokens per second (TPS). Each model's estimated IQ is plotted against how quickly it streams its answer after generation begins.Data: AI IQ methodology, Artificial AnalysisOpen chartResponse TimeIQ vs LatencyLatency is time to first token (TTFT): how long a model takes to begin responding. Drag the input-length slider from 1K to 10K tokens to interpolate between AA's two measured TTFT anchors.Data: AI IQ methodology, Artificial AnalysisOpen chartResponse TimeIQ vs Completion TimeCompletion time is time to last token (TTLT): time to first answer token plus output tokens ÷ throughput (TPS). Drag the input slider (1K–10K tokens) and output slider (100–10K tokens) to see how TTLT changes with the workload.Data: AI IQ methodology, Artificial AnalysisOpen chartResponse TimeIQ vs Completion Time vs Cost in 3D3D scatter: X = completion time (time to last token, TTLT; log, faster to the right), Y = IQ, Z = effective cost (log). Color = provider. Drag to rotate.Data: Artificial Analysis, ARC Prize, Vals.aiOpen chartResponse TimeAnimating Completion Time over IQ vs CostX = Effective Cost/1M I/O Tokens (log). Y = IQ. Clock spin rate represents completion time (time to last token, TTLT).Data: AI IQ methodology, Artificial AnalysisOpen chartCostIQ vs Effective CostEach model's estimated IQ plotted against effective cost per 1M I/O Tokens (sticker price × measured or imputed usage multiplier).Data: AI IQ methodologyOpen chart