AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Gemini 3.6 FlashGPT-5.3-Codex
Gemini 3.6 Flash

Google

GPT-5.3-Codex

OpenAI

Pricing
Input $/1M$1.50$1.75
Output $/1M$7.50$14.00
Blended $/1M (3:1)$3.00$4.81
Cache read $/1M$0.15$0.175
Cache write $/1M$0.083
Quality indices
Intelligence Index50.144.3
Coding Index69.2
Math Index
Agentic Index38.7
Speed & latency
Output speed (tok/s)311126
Time to first token12.80s51.38s
Time to first answer token12.80s51.38s
Benchmarks (Artificial Analysis)
GPQA Diamond92.8%91.5%
Humanity's Last Exam38.3%39.9%
SciCode52.7%53.2%
IFBench75.4%
AA-LCR (Long Context Reasoning)69.7%74.0%
Terminal-Bench Hard53.0%
Terminal-Bench 2.177.5%
τ²-Bench (Telecom)86.0%
τ³-Bench Banking24.5%
Design Arena (Elo)
Website1,3321,190
Web apps1,103
Full-stack1,041
Mobile apps1,127
Android native1,084
UI components1,179
Data viz1,199
SVG1,176
3D1,067
Game dev1,219
Godot game dev1,124
Code categories1,178
ASCII art1,193
Specs
Context window1.05M400K
Max output tokens66K128K
Input modalitiestext, image, video, file, audiotext, image, file
TokenizerGeminiGPT