AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

GPT-5.6 SolGrok 4.3 (batch)
GPT-5.6 Sol

OpenAI

Grok 4.3 (batch)

xAI

Pricing
Input $/1M$2.00$1.00
Output $/1M$10.00$2.00
Blended $/1M (3:1)$4.00$1.25
Cache read $/1M$0.20$0.16
Cache write $/1M$2.50
Quality indices
Intelligence Index4724.9
Coding Index77.442.2
Math Index
Agentic Index50.215.5
Speed & latency
Output speed (tok/s)650
Time to first token64.46s0s
Time to first answer token64.46s0s
Benchmarks (Artificial Analysis)
GPQA Diamond94.1%90.1%
Humanity's Last Exam49.5%37.2%
SciCode57.1%48.3%
IFBench72.7%81.3%
AA-LCR (Long Context Reasoning)84.0%73.0%
Terminal-Bench Hard65.9%37.9%
Terminal-Bench 2.188.0%39.7%
τ²-Bench (Telecom)85.1%97.7%
τ³-Bench Banking44.3%12.4%
Design Arena (Elo)
Website1,207
Web apps1,144
Full-stack1,017
Mobile apps1,092
Android native972
UI components1,207
Data viz1,197
SVG1,109
3D1,154
Game dev1,198
Agentic game dev1,007
Godot game dev1,035
HTML slides1,011
PPTX slides1,071
Python→PPTX slides1,072
Agentic slides1,072
Agentic HTML slides1,066
Agentic slides (HTML)1,066
Agentic slides (Python PPTX)1,068
Code categories1,202
ASCII art1,157
Specs
Context window1.05M1M
Max output tokens128K900K
Input modalitiesfile, image, texttext, image, file
TokenizerGPTGrok