← All models

Anthropic

glm-5-turbo

z-ai/glm-5-turbo

Price per million tokens

Input
$1.519
Output
$5.064
Cache write
not offered
Cache read
$0.304

What you pay. It is what the tokens cost us to buy plus 20%, and your usage page shows what we paid beside what you paid on every request you send.

Limits

Context window
202,752 tokens
Max output
131,072 tokens
Images
text only

What it costs to run

WorkloadInputOutputCostPer 1,000 calls
Short chat turn1,000500$0.0041$4.05
Long document summary50,0002,000$0.0861$86.09
Agent step with tools20,0001,500$0.0380$37.98
1M in, 1M out1,000,0001,000,000$6.5832$6583.20

Calling it

const res = await client.chat.completions.create({
  model: "z-ai/glm-5-turbo",
  messages: [{ role: "user", content: "Hello" }],
});