AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Gemini 3.1 Pro Preview (batch)GPT-5.6 Sol
Gemini 3.1 Pro Preview (batch)

Google

GPT-5.6 Sol

OpenAI

Pricing
Input $/1M$1.00$2.00
Output $/1M$6.00$10.00
Blended $/1M (3:1)$2.25$4.00
Cache read $/1M$0.20
Cache write $/1M$2.50
Quality indices
Intelligence Index29.747
Coding Index68.877.4
Math Index
Agentic Index8.250.2
Speed & latency
Output speed (tok/s)13265
Time to first token24.64s64.46s
Time to first answer token24.64s64.46s
Benchmarks (Artificial Analysis)
GPQA Diamond94.1%94.1%
Humanity's Last Exam47.0%49.5%
SciCode58.7%57.1%
IFBench77.1%72.7%
AA-LCR (Long Context Reasoning)82.0%84.0%
Terminal-Bench Hard53.8%65.9%
Terminal-Bench 2.173.8%88.0%
τ²-Bench (Telecom)95.6%85.1%
τ³-Bench Banking21.4%44.3%
Design Arena (Elo)
Website1,265
Web apps1,141
Full-stack1,078
Mobile apps1,121
Android native1,063
UI components1,295
Data viz1,254
SVG1,309
3D1,263
Game dev1,226
Agentic game dev1,114
Godot game dev1,236
HTML slides1,155
PPTX slides1,110
Python→PPTX slides1,109
Agentic slides1,112
Agentic HTML slides1,226
Agentic slides (HTML)1,219
Agentic slides (Python PPTX)1,107
Code categories1,259
ASCII art1,289
Specs
Context window1.05M1.05M
Max output tokens66K128K
Input modalitiesaudio, file, image, text, videofile, image, text
TokenizerGeminiGPT