Nemotron 3 Super 120B A12B

Nvidia llm

Mid-size Nemotron 3 model using a latent mixture-of-experts hybrid with multi-token prediction, targeting multi-agent applications over a 1M-token context.

overall ⌄

38.1

62 of 93 ranked

price · $/M tokens

self-hosted

open weights, no first-party API

to run it yourself ⌄

72 GB

minimum · 288 GB recommended

Benchmarks

What its maker published, and what anyone else measured. Every figure links the document it came from.

benchmarkpublishedmeasured
MMLU-Pro83.73
AIME 202590.21
SWE-bench60.47
HLE18.26
BrowseComp31.28

Lineage

No recorded lineage — root or standalone model.

view in full graph →

Metadata

orgNvidia
released2026-03-11
params120B-A12B
licensenvidia-open-model
hf idnvidia/NVIDIA-Nemotron-3-Super-120B-A12B-BF16