models · nvidia

Llama-3.3-Nemotron-Super-49B-v1

NVIDIA pruned + distilled 49B from Llama-3.3-70B via NAS. Near-70B quality at 49B.

Compare Verify

What this registry has checked

Quarantine. We hold the publisher’s own published hash for these weights and have recorded it against this entry. We have not downloaded and hashed the weights ourselves, so the hash is the publisher’s claim, not our measurement.
Weight file SHA-256957f16be72a083739afe6de9d35856ef8a80978d098701b2ad76d30c6d57cdb7
Methodhf-published-sha256

What is recorded

Publishernvidia
OriginUnited States
LicenceLlama-3.3
Publisher’s own licence statementother
Parameters49,867,145,216 (49.9B)
Architecturenemotron-nas
Layers80
Hidden size8192
AttentionMHA (64 heads, 64 KV)
Native precisionBF16
Pulls recorded here42,014
Downloads reported upstream101,756

Reported scores

MMLU-Pro69.8
GPQA Diamond51.7
AIME 20257.7
MATH-50077.5
LiveCodeBench28
IFBench39.5
Humanity's Last Exam3.8

Reported by a third party and recorded here, not re-run by us. Source: artificialanalysis.ai/api/v2.

Elsewhere in the registry

Registry record as of 2026-06-11. Identity, licence and parameter count are properties of a release and do not move; tier can, and the verify link above re-checks it against the live registry rather than this page.