AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

R1GPT-5.6 Sol
R1

DeepSeek

GPT-5.6 Sol

OpenAI

Pricing
Input $/1M$0.70$5.00
Output $/1M$2.50$30.00
Blended $/1M (3:1)$1.15$11.25
Cache read $/1M$0.50
Cache write $/1M$6.25
Quality indices
Intelligence Index20.158.9
Coding Index24.677.4
Math Index76
Agentic Index3.154
Speed & latency
Output speed (tok/s)067
Time to first token0s120.38s
Time to first answer token0s120.38s
Benchmarks (Artificial Analysis)
MMLU-Pro84.9%
GPQA Diamond81.3%94.1%
Humanity's Last Exam14.9%47.2%
LiveCodeBench77.0%
SciCode40.3%56.1%
MATH-50098.3%
AIME89.3%
AIME 202576.0%
IFBench39.6%72.7%
AA-LCR (Long Context Reasoning)54.7%73.7%
Terminal-Bench Hard15.9%65.9%
Terminal-Bench 2.188.0%
τ²-Bench (Telecom)36.5%85.1%
τ³-Bench Banking33.0%
Specs
Context window164K1.05M
Max output tokens16K128K
Input modalitiestextfile, image, text
TokenizerDeepSeekGPT