AiCostCompare

GLM 5V Turbo

by Z.ai · z-ai/glm-5v-turbo · #292 cheapest of 422 paid models

Prices updated Sep 19, 2026, 5:37 PM UTC · refreshed hourly

GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...

Compare vs:Claude Sonnet 5Claude Opus 4.6Claude Haiku 4.5GPT-5.6 Sol

Pricing

Input

$1.20

per 1M tokens

Output

$4.00

per 1M tokens

Blended (3:1)

$1.90

3 input : 1 output

Cache read

$0.24

per 1M cached tokens

Cache write

per 1M tokens

Cache & batch economics

Effective prices for GLM 5V Turbo with prompt caching.

Cache hit rateEffective blended $/1MExample request*vs no cache
0%$1.90$0.1264
50%$1.54$0.0784−38%
90%$1.25$0.04−68%

*Example: 100K sticky context + 2K new input + 1K output. Batch discount is an approximate 50% off list rates (available for OpenAI / Anthropic / Google — not flagged for this creator). Confirm on the vendor’s pricing page.

What a request costs

WorkloadInput tokensOutput tokensCost / requestCost / 1K requests
Short chat message500300$0.0018$1.80
Document summary8,0001,000$0.0136$13.60
Codebase question (RAG)30,0002,000$0.044$44.00
Long-context analysis150,0005,000$0.2$200.00

Performance

Intelligence Index

AA composite quality

Coding Index

Math Index

Agentic Index

Output speed

median

Time to first token

median

Time to first answer

after reasoning tokens

Design Arena Elo

CategoryElo
Website1,243
Web apps1,139
Full-stack1,159
Mobile apps1,163
Android native1,267
UI components1,228
Data viz1,215
SVG1,174
3D1,238
Game dev1,241
Agentic game dev1,102
Godot game dev970
HTML slides1,123
PPTX slides1,164
Python→PPTX slides1,165
Agentic slides1,171
Agentic HTML slides1,134
Agentic slides (HTML)1,137
Agentic slides (Python PPTX)1,183
Code categories1,244
ASCII art1,131

Specs

Context window

203K

tokens

Max output

131K

tokens

Input modalities

image, text, video

Tokenizer

Other