Nemotron 3.5 Lightning pricing
$$ Standard$0.08/M input · $0.20/M output. List prices from provider data; in idapt, chat is metered per token with transparent, usage-based pricing at the prices shown in the app.
Full rate table
| Rate | Price | Unit |
|---|---|---|
| Input | 0.08 | /M tokens |
| Output | 0.19999999999999998 | /M tokens |
| Cache read | 0.04 | /M tokens |
What typical work costs
| Workload | Tokens | List price |
|---|---|---|
| Quick question | 1K tokens in, 500 out | $0.0002 |
| Long-document summary | 60K tokens in, 2K out | $0.0052 |
| Agent working session | 400K tokens in, 40K out across a day | $0.04 |
Price your own workload
- Per call
- $0.0002
- Per month
- $0.02
Compare up to four models at once in the cost calculator.
Frequently asked
What does Nemotron 3.5 Lightning cost per 1M tokens?
Nemotron 3.5 Lightning lists at $0.08 per 1M input tokens and $0.20 per 1M output tokens.
Does Nemotron 3.5 Lightning discount cached input tokens?
Yes. Cached input tokens are billed at $0.04 per 1M, versus $0.08 for fresh input. Prompts that repeat a long prefix (a system prompt, attached files) benefit the most.
What does a typical chat message with Nemotron 3.5 Lightning cost?
A message with about 1,000 input tokens and a 500-token reply costs roughly $0.0002 at list price. Long documents and long replies scale that linearly with token counts.
Is Nemotron 3.5 Lightning cheaper than Nemotron 3 Ultra?
Nemotron 3.5 Lightning has the lower blended price at a 3:1 input:output mix. Nemotron 3.5 Lightning: $0.08 in / $0.20 out per 1M. Nemotron 3 Ultra: $0.60 in / $3.60 out per 1M.
How does idapt bill Nemotron 3.5 Lightning?
Chat with Nemotron 3.5 Lightning in idapt is metered per token with transparent, usage-based pricing at the prices shown in the app. Subscriptions include a weekly usage allowance before any balance is touched.
Part of the Nemotron 3.5 Lightning model page · run it locally · cheapest models board · price changes