AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

Gemini 3.8 Flash (batch)GPT-5.6 Sol
Gemini 3.8 Flash (batch)

Google

GPT-5.6 Sol

OpenAI

Pricing
Input $/1M$0.375$2.00
Output $/1M$1.88$10.00
Blended $/1M (3:1)$0.75$4.00
Cache read $/1M$0.037$0.20
Cache write $/1M$0.042$2.50
Quality indices
Intelligence Index40.947
Coding Index76.377.4
Math Index
Agentic Index40.250.2
Speed & latency
Output speed (tok/s)32765
Time to first token12.22s64.46s
Time to first answer token12.22s64.46s
Benchmarks (Artificial Analysis)
GPQA Diamond95.3%94.1%
Humanity's Last Exam47.8%49.5%
SciCode56.6%57.1%
IFBench72.7%
AA-LCR (Long Context Reasoning)81.3%84.0%
Terminal-Bench Hard65.9%
Terminal-Bench 2.187.6%88.0%
τ²-Bench (Telecom)85.1%
τ³-Bench Banking44.9%44.3%
Design Arena (Elo)
Website1,315
Web apps1,256
Full-stack1,250
Mobile apps1,259
UI components1,338
Data viz1,264
3D1,324
Game dev1,340
Python→PPTX slides1,176
Code categories1,323
Specs
Context window1.05M1.05M
Max output tokens66K128K
Input modalitiestext, image, video, file, audiofile, image, text
TokenizerGeminiGPT