AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Gemini 3.6 FlashGPT-5.1
Gemini 3.6 Flash

Google

GPT-5.1

OpenAI

Pricing
Input $/1M$1.50$1.25
Output $/1M$7.50$10.00
Blended $/1M (3:1)$3.00$3.44
Cache read $/1M$0.15$0.125
Cache write $/1M$0.083
Quality indices
Intelligence Index50.136.9
Coding Index69.249.4
Math Index94
Agentic Index38.721
Speed & latency
Output speed (tok/s)301110
Time to first token11.80s39.50s
Time to first answer token11.80s39.50s
Benchmarks (Artificial Analysis)
MMLU-Pro87.0%
GPQA Diamond92.8%87.3%
Humanity's Last Exam38.3%26.5%
LiveCodeBench86.8%
SciCode52.7%43.3%
AIME 202594.0%
IFBench72.9%
AA-LCR (Long Context Reasoning)69.7%75.0%
Terminal-Bench Hard45.5%
Terminal-Bench 2.177.5%52.4%
τ²-Bench (Telecom)81.9%
τ³-Bench Banking24.5%14.0%
Design Arena (Elo)
Website1,3321,215
Web apps1,075
Mobile apps1,122
UI components1,209
Data viz1,241
SVG1,196
3D1,120
Game dev1,237
Code categories1,204
ASCII art1,160
Specs
Context window1.05M400K
Max output tokens66K128K
Input modalitiestext, image, video, file, audioimage, text, file
TokenizerGeminiGPT