AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Gemini 3.6 FlashGPT-5.5 (batch)
Gemini 3.6 Flash

Google

GPT-5.5 (batch)

OpenAI

Pricing
Input $/1M$0.75$2.50
Output $/1M$3.75$15.00
Blended $/1M (3:1)$1.50$5.63
Cache read $/1M$0.075$0.25
Cache write $/1M$0.042
Quality indices
Intelligence Index3438.4
Coding Index69.274.9
Math Index
Agentic Index2936.4
Speed & latency
Output speed (tok/s)2650
Time to first token21.43s0s
Time to first answer token21.43s0s
Benchmarks (Artificial Analysis)
GPQA Diamond92.8%93.5%
Humanity's Last Exam40.8%45.8%
SciCode53.4%55.8%
IFBench75.9%
AA-LCR (Long Context Reasoning)80.0%84.3%
Terminal-Bench Hard60.6%
Terminal-Bench 2.177.5%84.3%
τ²-Bench (Telecom)93.9%
τ³-Bench Banking29.9%39.0%
Design Arena (Elo)
Website1,3121,269
Web apps1,2111,132
Full-stack1,1881,099
Mobile apps1,2281,171
Android native1,2131,175
UI components1,3181,273
Data viz1,3111,273
SVG1,260
3D1,3021,226
Game dev1,2861,320
Agentic game dev1,1801,179
Godot game dev1,183
HTML slides1,1521,070
PPTX slides1,157
Python→PPTX slides1,1461,152
Agentic slides1,150
Agentic HTML slides1,084
Agentic slides (HTML)1,077
Agentic slides (Python PPTX)1,155
Code categories1,3051,274
ASCII art1,3041,276
Specs
Context window1.05M1.05M
Max output tokens66K128K
Input modalitiestext, image, video, file, audiofile, image, text
TokenizerGeminiGPT