NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...
Specifications
| Provider | NVIDIA |
|---|---|
| Context window | 1,000,000 tokens (1M) |
| Input price | $0.085 / 1M tokens |
| Output price | $0.4 / 1M tokens |
| Input modalities | text |
| Output modalities | text |
| Model ID | nvidia/nemotron-3-super-120b-a12b |
Pricing is per 1M tokens, indicative pricing via OpenRouter (updated 2026-08-22); a provider's own list price may differ.
Frequently asked questions
How much does Nemotron 3 Super cost?
Indicative pricing (via OpenRouter, updated 2026-08-22) is $0.085 per 1M input tokens and $0.4 per 1M output tokens. A provider's own list price may differ. On Vincony, Nemotron 3 Super is available on credit-based pricing alongside 800+ other models.
What is Nemotron 3 Super's context window?
Nemotron 3 Super supports a context window of up to 1,000,000 tokens (1M), which is how much text it can consider at once.
Who makes Nemotron 3 Super?
Nemotron 3 Super is built by NVIDIA. You can access it — and compare it against other providers' models — through Vincony's multi-model platform.
How do I use Nemotron 3 Super?
You can use Nemotron 3 Super directly through NVIDIA, or via Vincony, which aggregates 800+ models (including Nemotron 3 Super) into one interface with side-by-side comparison and credit-based pricing.