AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Gemma 3 12BGPT-5.6 Sol
Gemma 3 12B

Google

GPT-5.6 Sol

OpenAI

Pricing
Input $/1M$0.05$5.00
Output $/1M$0.15$30.00
Blended $/1M (3:1)$0.075$11.25
Cache read $/1M$0.50
Cache write $/1M$6.25
Quality indices
Intelligence Index5.558.9
Coding Index5.877.4
Math Index18.3
Agentic Index0.354
Speed & latency
Output speed (tok/s)070
Time to first token0s112.90s
Time to first answer token0s112.90s
Benchmarks (Artificial Analysis)
MMLU-Pro59.5%
GPQA Diamond34.9%94.1%
Humanity's Last Exam4.8%47.2%
LiveCodeBench13.7%
SciCode17.4%56.1%
MATH-50085.3%
AIME22.0%
AIME 202518.3%
IFBench36.7%72.7%
AA-LCR (Long Context Reasoning)6.7%73.7%
Terminal-Bench Hard0.8%65.9%
Terminal-Bench 2.10.0%88.0%
τ²-Bench (Telecom)10.8%85.1%
τ³-Bench Banking0.8%33.0%
Specs
Context window131K1.05M
Max output tokens16K128K
Input modalitiestext, imagefile, image, text
TokenizerGeminiGPT