Nemotron 3 Ultra

NVIDIA’s largest open model: hybrid Mamba-Transformer MoE, 55B active.

Nemotron 3 Ultra is NVIDIA’s open frontier-reasoning and orchestration model, with 55B active out of 550B total parameters, built on a hybrid Transformer-Mamba mixture-of-experts architecture.

Provider
NVIDIA
Type
Language model
Released
Jun 4, 2026
Context window
262K tokens
Max output
16K tokens
Price
$0.50 input / $2.20 output per 1M tokens
Input
text
Output
text
Open weights
Yes
License
NVIDIA Open Model License
Knowledge cutoff
Sep 30, 2025

Best for

Strengths

More from NVIDIA

Tools that use it

Sources