AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Llama 4 ScoutGPT-5.6 Sol
Llama 4 Scout

Meta

GPT-5.6 Sol

OpenAI

Pricing
Input $/1M$0.10$5.00
Output $/1M$0.30$30.00
Blended $/1M (3:1)$0.15$11.25
Cache read $/1M$0.50
Cache write $/1M$6.25
Quality indices
Intelligence Index1058.9
Coding Index8.277.4
Math Index14
Agentic Index1.154
Speed & latency
Output speed (tok/s)7667
Time to first token0.56s120.38s
Time to first answer token0.56s120.38s
Benchmarks (Artificial Analysis)
MMLU-Pro75.2%
GPQA Diamond58.7%94.1%
Humanity's Last Exam4.3%47.2%
LiveCodeBench29.9%
SciCode17.0%56.1%
MATH-50084.4%
AIME28.3%
AIME 202514.0%
IFBench39.5%72.7%
AA-LCR (Long Context Reasoning)25.8%73.7%
Terminal-Bench Hard1.5%65.9%
Terminal-Bench 2.13.7%88.0%
τ²-Bench (Telecom)15.5%85.1%
τ³-Bench Banking3.3%33.0%
Design Arena (Elo)
Website778
UI components811
Data viz933
Game dev831
Code categories823
Specs
Context window1.31M1.05M
Max output tokens16K128K
Input modalitiestext, imagefile, image, text
TokenizerLlama4GPT