AiCostCompare

Compare models

Pick up to 4 models for a side-by-side spec sheet. Hover any dotted metric name for what the test measures. Best value in each row is highlighted.

GPT-5.6 SolQwen3.7 Max
GPT-5.6 Sol

OpenAI

Qwen3.7 Max

Alibaba (Qwen)

Pricing
Input $/1M$5.00$1.48
Output $/1M$30.00$4.42
Blended $/1M (3:1)$11.25$2.21
Cache read $/1M$0.50$0.295
Cache write $/1M$6.25$1.84
Quality indices
Intelligence Index58.946
Coding Index77.466
Math Index
Agentic Index5430.6
Speed & latency
Output speed (tok/s)67207
Time to first token120.38s1.55s
Time to first answer token120.38s13.16s
Benchmarks (Artificial Analysis)
GPQA Diamond94.1%92.3%
Humanity's Last Exam47.2%38.1%
SciCode56.1%48.8%
IFBench72.7%80.5%
AA-LCR (Long Context Reasoning)73.7%69.0%
Terminal-Bench Hard65.9%50.8%
Terminal-Bench 2.188.0%74.5%
τ²-Bench (Telecom)85.1%94.7%
τ³-Bench Banking33.0%10.9%
Design Arena (Elo)
Website1,291
Web apps1,253
Full-stack1,225
Mobile apps1,208
Android native1,182
UI components1,313
Data viz1,304
SVG1,267
3D1,310
Game dev1,315
Agentic game dev1,180
Godot game dev1,248
HTML slides1,191
Python→PPTX slides1,229
Code categories1,299
ASCII art1,261
Specs
Context window1.05M1M
Max output tokens128K66K
Input modalitiesfile, image, texttext
TokenizerGPTQwen