Home / Models / Nemotron 3 Super 120B-A12B
Nemotron 3 Super 120B-A12B
Maker: NVIDIA · reasoning
Open hybrid Mamba-2/Transformer sparse-MoE model: 120B total, 12B active, native 1M-token context, pretrained on 25T tokens in NVFP4. Over 5x the throughput of the previous Nemotron Super and 85.6% on PinchBench. NVIDIA Nemotron Open Model License.
Understands
text
This model cannot see images or video — you can only give it text. That matters if you planned to show it screenshots or documents with pictures.
Produces
text, code
Contents verified with the vendor: 2026-08-18
Included in subscriptions
No catalogue entry marks this model as part of a plan yet. It will appear here once the plan contents are confirmed with the vendor.