AiCostCompare

GLM 5V Turbo

by Z.ai · z-ai/glm-5v-turbo · #220 cheapest of 325 paid models

Prices updated Jul 22, 2026, 1:43 AM UTC · refreshed hourly

GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...

Compare vs:Claude Sonnet 5Claude Opus 4.6Claude Haiku 4.5GPT-5.6 Sol

Pricing

Input

$1.20

per 1M tokens

Output

$4.00

per 1M tokens

Blended (3:1)

$1.90

3 input : 1 output

Cache read

$0.24

per 1M cached tokens

Cache write

per 1M tokens

Cache & batch economics

Effective prices for GLM 5V Turbo with prompt caching.

Cache hit rateEffective blended $/1MExample request*vs no cache
0%$1.90$0.1264
50%$1.54$0.0784−38%
90%$1.25$0.04−68%

*Example: 100K sticky context + 2K new input + 1K output. Batch discount is an approximate 50% off list rates (available for OpenAI / Anthropic / Google — not flagged for this creator). Confirm on the vendor’s pricing page.

What a request costs

WorkloadInput tokensOutput tokensCost / requestCost / 1K requests
Short chat message500300$0.0018$1.80
Document summary8,0001,000$0.0136$13.60
Codebase question (RAG)30,0002,000$0.044$44.00
Long-context analysis150,0005,000$0.2$200.00

Performance

Intelligence Index

AA composite quality

Coding Index

Math Index

Agentic Index

Output speed

median

Time to first token

median

Time to first answer

after reasoning tokens

Design Arena Elo

CategoryElo
Website1,256
Web apps1,168
Full-stack1,201
Mobile apps1,205
Android native1,267
UI components1,253
Data viz1,237
SVG1,198
3D1,269
Game dev1,280
Agentic game dev1,105
Godot game dev1,177
HTML slides1,135
PPTX slides1,164
Python→PPTX slides1,165
Agentic slides1,171
Agentic HTML slides1,134
Agentic slides (HTML)1,137
Agentic slides (Python PPTX)1,183
Code categories1,262
ASCII art1,142

Specs

Context window

203K

tokens

Max output

131K

tokens

Input modalities

image, text, video

Tokenizer

Other