AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

DeepSeek V3.1 TerminusGPT-5.6 Sol
DeepSeek V3.1 Terminus

DeepSeek

GPT-5.6 Sol

OpenAI

Pricing
Input $/1M$0.27$5.00
Output $/1M$1.00$30.00
Blended $/1M (3:1)$0.453$11.25
Cache read $/1M$0.135$0.50
Cache write $/1M$6.25
Quality indices
Intelligence Index21.458.9
Coding Index43.577.4
Math Index53.7
Agentic Index18.154
Speed & latency
Output speed (tok/s)067
Time to first token0s120.38s
Time to first answer token0s120.38s
Benchmarks (Artificial Analysis)
MMLU-Pro83.6%
GPQA Diamond75.1%94.1%
Humanity's Last Exam8.4%47.2%
LiveCodeBench52.9%
SciCode32.1%56.1%
AIME 202553.7%
IFBench41.2%72.7%
AA-LCR (Long Context Reasoning)43.3%73.7%
Terminal-Bench Hard31.8%65.9%
Terminal-Bench 2.188.0%
τ²-Bench (Telecom)37.1%85.1%
τ³-Bench Banking33.0%
Design Arena (Elo)
Website1,214
UI components1,224
Data viz1,200
SVG1,112
3D1,199
Game dev1,189
Code categories1,209
Specs
Context window164K1.05M
Max output tokens33K128K
Input modalitiestextfile, image, text
TokenizerDeepSeekGPT