AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Gemma 3 4BGPT-5.6 Sol
Gemma 3 4B

Google

GPT-5.6 Sol

OpenAI

Pricing
Input $/1M$0.05$5.00
Output $/1M$0.10$30.00
Blended $/1M (3:1)$0.063$11.25
Cache read $/1M$0.50
Cache write $/1M$6.25
Quality indices
Intelligence Index1.158.9
Coding Index2.777.4
Math Index12.7
Agentic Index54
Speed & latency
Output speed (tok/s)067
Time to first token0s120.38s
Time to first answer token0s120.38s
Benchmarks (Artificial Analysis)
MMLU-Pro41.7%
GPQA Diamond29.1%94.1%
Humanity's Last Exam5.2%47.2%
LiveCodeBench11.2%
SciCode7.3%56.1%
MATH-50076.6%
AIME6.3%
AIME 202512.7%
IFBench28.3%72.7%
AA-LCR (Long Context Reasoning)5.7%73.7%
Terminal-Bench Hard0.8%65.9%
Terminal-Bench 2.10.4%88.0%
τ²-Bench (Telecom)5.0%85.1%
τ³-Bench Banking0.4%33.0%
Specs
Context window131K1.05M
Max output tokens16K128K
Input modalitiestext, imagefile, image, text
TokenizerGeminiGPT