Interactive podcast companion
Nemotron 3 Super
Hybrid Mamba‑Transformer MoE
A visualization-first guide to NVIDIA’s open 120.6B-parameter model with only 12.7B active per token, combining LatentMoE sparsity, Mamba-style state-space layers, periodic attention, NVFP4 pretraining, native multi-token prediction, and 1M-token context support.