← All models

Anthropic

nemotron-3.5-lightning

nvidia/nemotron-3.5-lightning

Price per million tokens

Input
$0.101
Output
$0.253
Cache write
not offered
Cache read
$0.051

What you pay. It is what the tokens cost us to buy plus 20%, and your usage page shows what we paid beside what you paid on every request you send.

Limits

Context window
262,144 tokens
Max output
131,072 tokens
Images
text only

What it costs to run

WorkloadInputOutputCostPer 1,000 calls
Short chat turn1,000500$0.0002$0.23
Long document summary50,0002,000$0.0056$5.57
Agent step with tools20,0001,500$0.0024$2.41
1M in, 1M out1,000,0001,000,000$0.3545$354.48

Calling it

const res = await client.chat.completions.create({
  model: "nvidia/nemotron-3.5-lightning",
  messages: [{ role: "user", content: "Hello" }],
});