← All models

Anthropic

glm-5.3-flash

z-ai/glm-5.3-flash

Price per million tokens

Input
$0.089
Output
$0.295
Cache write
not offered
Cache read
$0.018

What you pay, and your usage page shows what we paid beside what you paid on every request you send.

Limits

Context window
1,310,720 tokens
Max output
131,072 tokens
Images
supported

What it costs to run

WorkloadInputOutputCostPer 1,000 calls
Short chat turn1,000500$0.0002$0.24
Long document summary50,0002,000$0.0050$5.02
Agent step with tools20,0001,500$0.0022$2.22
1M in, 1M out1,000,0001,000,000$0.3840$384.02

Calling it

const res = await client.chat.completions.create({
  model: "z-ai/glm-5.3-flash",
  messages: [{ role: "user", content: "Hello" }],
});