NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...
Specifications
| Provider | NVIDIA |
|---|---|
| Context window | 262,144 tokens (262K) |
| Input price | $0.08 / 1M tokens |
| Output price | $0.2 / 1M tokens |
| Input modalities | text |
| Output modalities | text |
| Model ID | nvidia/nemotron-3.5-lightning |
Pricing is per 1M tokens, indicative pricing via OpenRouter (updated 2026-08-22); a provider's own list price may differ.
Frequently asked questions
How much does Nemotron 3.5 Lightning cost?
Indicative pricing (via OpenRouter, updated 2026-08-22) is $0.08 per 1M input tokens and $0.2 per 1M output tokens. A provider's own list price may differ. On Vincony, Nemotron 3.5 Lightning is available on credit-based pricing alongside 800+ other models.
What is Nemotron 3.5 Lightning's context window?
Nemotron 3.5 Lightning supports a context window of up to 262,144 tokens (262K), which is how much text it can consider at once.
Who makes Nemotron 3.5 Lightning?
Nemotron 3.5 Lightning is built by NVIDIA. You can access it — and compare it against other providers' models — through Vincony's multi-model platform.
How do I use Nemotron 3.5 Lightning?
You can use Nemotron 3.5 Lightning directly through NVIDIA, or via Vincony, which aggregates 800+ models (including Nemotron 3.5 Lightning) into one interface with side-by-side comparison and credit-based pricing.