Nemotron 3 Super
Open hybrid Mamba-Transformer MoE for multi-agent applications.
NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model activating 12B parameters for compute efficiency in complex multi-agent applications.
- Provider
- NVIDIA
- Type
- Language model
- Released
- Mar 11, 2026
- Context window
- 262K tokens
- Max output
- 236K tokens
- Price
- $0.080 input / $0.45 output per 1M tokens
- Input
- text
- Output
- text
- Open weights
- Yes
- License
- NVIDIA Open Model License
- Knowledge cutoff
- Jun 1, 2025
Best for
- Agents
- Local & on-device
- Enterprise
- Real-time / low latency
Strengths
- Hybrid Mamba-Transformer
- Optimized for NVIDIA GPUs
More from NVIDIA
- Nemotron 3.5 Lightning — NVIDIA. Tiny open 30B-A3B model with the lowest first-token latency in the index.
- Nemotron 3 Ultra — NVIDIA. NVIDIA’s largest open model: hybrid Mamba-Transformer MoE, 55B active.
- Nemotron 3 Embed 1B — NVIDIA. Small open embedding model for high-throughput enterprise retrieval.
- Cosmos 3 — NVIDIA. NVIDIA’s open omni-modal world foundation model for physical AI.
- Isaac GR00T N1.7 — NVIDIA. Open 3B cross-embodiment VLA for humanoids and manipulators.
- Nemotron 3.5 Content Safety — NVIDIA. Compact 4B multimodal guardrail model for LLM and VLM traffic.