Home / Models / Nemotron 3 Super 120B-A12B

Nemotron 3 Super 120B-A12B

Maker: NVIDIA · reasoning

Open hybrid Mamba-2/Transformer sparse-MoE model: 120B total, 12B active, native 1M-token context, pretrained on 25T tokens in NVFP4. Over 5x the throughput of the previous Nemotron Super and 85.6% on PinchBench. NVIDIA Nemotron Open Model License.

Understands

text

This model cannot see images or video — you can only give it text. That matters if you planned to show it screenshots or documents with pictures.

Produces

text, code

Contents verified with the vendor: 2026-08-18

Included in subscriptions

No catalogue entry marks this model as part of a plan yet. It will appear here once the plan contents are confirmed with the vendor.