AiCostCompare

GLM 5.2 (free)

FREE

by Z.ai · z-ai/glm-5.2:free

Prices updated Sep 20, 2026, 12:34 AM UTC · refreshed hourly

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

Compare vs:Claude Sonnet 5Claude Opus 4.6Claude Haiku 4.5GPT-5.6 Sol

Pricing

Input

Free

per 1M tokens

Output

Free

per 1M tokens

Blended (3:1)

Free

3 input : 1 output

Cache read

per 1M cached tokens

Cache write

per 1M tokens

Cache & batch economics

Effective prices for GLM 5.2 (free) (no cache-read rate published — hits billed as normal input).

Cache hit rateEffective blended $/1MExample request*vs no cache
0%Free$0
50%Free$0−0%
90%Free$0−0%

*Example: 100K sticky context + 2K new input + 1K output. Batch discount is an approximate 50% off list rates (available for OpenAI / Anthropic / Google — not flagged for this creator). Confirm on the vendor’s pricing page.

What a request costs

WorkloadInput tokensOutput tokensCost / requestCost / 1K requests
Short chat message500300$0$0
Document summary8,0001,000$0$0
Codebase question (RAG)30,0002,000$0$0
Long-context analysis150,0005,000$0$0

Performance

Intelligence Index

34

AA composite quality

Coding Index

68.8

Math Index

Agentic Index

39.4

Output speed

median

Time to first token

median

Time to first answer

after reasoning tokens

Design Arena Elo

CategoryElo
Website1,304
Web apps1,236
Full-stack1,251
Mobile apps1,190
Android native1,189
UI components1,308
Data viz1,310
SVG1,247
3D1,321
Game dev1,296
Agentic game dev1,205
Godot game dev1,142
HTML slides1,178
Python→PPTX slides1,194
Code categories1,308
ASCII art1,229

Specs

Context window

33K

tokens

Max output

29K

tokens

Input modalities

text

Tokenizer

Other