AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

DeepSeek V3.1GPT-5.6 Sol
DeepSeek V3.1

DeepSeek

GPT-5.6 Sol

OpenAI

Pricing
Input $/1M$0.25$5.00
Output $/1M$0.95$30.00
Blended $/1M (3:1)$0.425$11.25
Cache read $/1M$0.13$0.50
Cache write $/1M$6.25
Quality indices
Intelligence Index2158.9
Coding Index77.4
Math Index49.7
Agentic Index54
Speed & latency
Output speed (tok/s)070
Time to first token0s112.90s
Time to first answer token0s112.90s
Benchmarks (Artificial Analysis)
MMLU-Pro83.3%
GPQA Diamond73.5%94.1%
Humanity's Last Exam6.3%47.2%
LiveCodeBench57.7%
SciCode36.7%56.1%
AIME 202549.7%
IFBench37.8%72.7%
AA-LCR (Long Context Reasoning)45.0%73.7%
Terminal-Bench Hard24.2%65.9%
Terminal-Bench 2.188.0%
τ²-Bench (Telecom)34.8%85.1%
τ³-Bench Banking33.0%
Design Arena (Elo)
Website1,150
UI components1,126
Data viz1,136
SVG1,014
3D1,136
Game dev1,143
Code categories1,145
Specs
Context window164K1.05M
Max output tokens33K128K
Input modalitiestextfile, image, text
TokenizerDeepSeekGPT