Model overview
- Model ID: nvidia/nemotron-3.5-lightning
- Provider: nvidia
- Type: language
- Context window: 262,144 tokens
- Max output tokens: 131,072
Tags: reasoning, tool-use, implicit-caching
Model pricing
| Metric | Value |
|---|---|
Input tokens (/1M) | $0.05 |
Output tokens (/1M) | $0.20 |
Image generation | n/a |
Cached input read (/1M) | $0.01 |
Cached input write (/1M) | n/a |
Pricing source | gateway |