Model overview
- Model ID: nvidia/nemotron-3.5-lightning-free
- Provider: nvidia
- Type: language
- Context window: 1,000,000 tokens
- Max output tokens: 32,768
Tags: reasoning, tool-use, implicit-caching, free
Model pricing
| Metric | Value |
|---|---|
Input tokens (/1M) | $0.00 |
Output tokens (/1M) | $0.00 |
Image generation | n/a |
Cached input read (/1M) | n/a |
Cached input write (/1M) | n/a |
Pricing source | gateway |