models · nvidia
NVIDIA pruned + distilled 49B from Llama-3.3-70B via NAS. Near-70B quality at 49B.
| Weight file SHA-256 | 957f16be72a083739afe6de9d35856ef8a80978d098701b2ad76d30c6d57cdb7 |
|---|---|
| Method | hf-published-sha256 |
| Publisher | nvidia |
|---|---|
| Origin | United States |
| Licence | Llama-3.3 |
| Publisher’s own licence statement | other |
| Parameters | 49,867,145,216 (49.9B) |
| Architecture | nemotron-nas |
| Layers | 80 |
| Hidden size | 8192 |
| Attention | MHA (64 heads, 64 KV) |
| Native precision | BF16 |
| Pulls recorded here | 42,014 |
| Downloads reported upstream | 101,756 |
| MMLU-Pro | 69.8 |
|---|---|
| GPQA Diamond | 51.7 |
| AIME 2025 | 7.7 |
| MATH-500 | 77.5 |
| LiveCodeBench | 28 |
| IFBench | 39.5 |
| Humanity's Last Exam | 3.8 |
Reported by a third party and recorded here, not re-run by us. Source: artificialanalysis.ai/api/v2.
Registry record as of 2026-06-11. Identity, licence and parameter count are properties of a release and do not move; tier can, and the verify link above re-checks it against the live registry rather than this page.