AiCostCompare

Qwen2.5 VL 72B Instruct

by Alibaba (Qwen) · qwen/qwen2.5-vl-72b-instruct · #169 cheapest of 325 paid models

Prices updated Jul 22, 2026, 2:33 AM UTC · refreshed hourly

Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects. It is also highly capable of analyzing texts, charts, icons, graphics, and layouts within images.

Compare vs:Claude Sonnet 5Claude Opus 4.6Claude Haiku 4.5GPT-5.6 Sol

Pricing

Input

$0.80

per 1M tokens

Output

$1.00

per 1M tokens

Blended (3:1)

$0.85

3 input : 1 output

Cache read

$0.40

per 1M cached tokens

Cache write

per 1M tokens

Cache & batch economics

Effective prices for Qwen2.5 VL 72B Instruct with prompt caching.

Cache hit rateEffective blended $/1MExample request*vs no cache
0%$0.85$0.0826
50%$0.70$0.0626−24%
90%$0.58$0.0466−44%

*Example: 100K sticky context + 2K new input + 1K output. Batch discount is an approximate 50% off list rates (available for OpenAI / Anthropic / Google — not flagged for this creator). Confirm on the vendor’s pricing page.

What a request costs

WorkloadInput tokensOutput tokensCost / requestCost / 1K requests
Short chat message500300$0.0007$0.7
Document summary8,0001,000$0.0074$7.40
Codebase question (RAG)30,0002,000$0.026$26.00
Long-context analysis150,0005,000$0.125$125.00

Specs

Context window

128K

tokens

Max output

128K

tokens

Input modalities

text, image

Tokenizer

Qwen