AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Llama 3.2 1B InstructGPT-5.6 Sol
Llama 3.2 1B Instruct

Meta

GPT-5.6 Sol

OpenAI

Pricing
Input $/1M$0.027$5.00
Output $/1M$0.201$30.00
Blended $/1M (3:1)$0.071$11.25
Cache read $/1M$0.50
Cache write $/1M$6.25
Quality indices
Intelligence Index58.9
Coding Index77.4
Math Index
Agentic Index54
Speed & latency
Output speed (tok/s)67
Time to first token120.38s
Time to first answer token120.38s
Benchmarks (Artificial Analysis)
GPQA Diamond94.1%
Humanity's Last Exam47.2%
SciCode56.1%
IFBench72.7%
AA-LCR (Long Context Reasoning)73.7%
Terminal-Bench Hard65.9%
Terminal-Bench 2.188.0%
τ²-Bench (Telecom)85.1%
τ³-Bench Banking33.0%
Specs
Context window60K1.05M
Max output tokens60K128K
Input modalitiestextfile, image, text
TokenizerLlama3GPT