Model Directory
    NVIDIA

    Nemotron 3 Ultra

    Intelligence 22.9Context 262KIn $0.5/MOut $2.2/M

    NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

    Specifications

    ProviderNVIDIA
    Intelligence Index22.9 (Artificial Analysis)
    Context window262,144 tokens (262K)
    Input price$0.5 / 1M tokens
    Output price$2.2 / 1M tokens
    Input modalitiestext
    Output modalitiestext
    Model IDnvidia/nemotron-3-ultra-550b-a55b

    Pricing is per 1M tokens, indicative pricing via OpenRouter (updated 2026-10-05); a provider's own list price may differ. Intelligence Index via Artificial Analysis.

    Frequently asked questions

    How capable is Nemotron 3 Ultra?

    Nemotron 3 Ultra scores 22.9 on the Artificial Analysis Intelligence Index, an aggregate of reasoning, knowledge, coding, and math benchmarks (higher is better). Compare it head-to-head with other models on the comparison pages below, or run it against them directly on Vincony.

    How much does Nemotron 3 Ultra cost?

    Indicative pricing (via OpenRouter, updated 2026-10-05) is $0.5 per 1M input tokens and $2.2 per 1M output tokens. A provider's own list price may differ. On Vincony, Nemotron 3 Ultra is available on credit-based pricing alongside 750+ other models.

    What is Nemotron 3 Ultra's context window?

    Nemotron 3 Ultra supports a context window of up to 262,144 tokens (262K), which is how much text it can consider at once.

    Who makes Nemotron 3 Ultra?

    Nemotron 3 Ultra is built by NVIDIA. You can access it — and compare it against other providers' models — through Vincony's multi-model platform.

    How do I use Nemotron 3 Ultra?

    You can use Nemotron 3 Ultra directly through NVIDIA, or via Vincony, which aggregates 750+ models (including Nemotron 3 Ultra) into one interface with side-by-side comparison and credit-based pricing.