models · nvidia
NVIDIA's 253B ultra model distilled from frontier models. Top open-weight reasoning.
| Weight file SHA-256 | 6d5c8ac78f27a83a8961236297c7b76cdfc84cfb4326e4ecae0db18b9e48a737 |
|---|---|
| Method | hf-published-sha256 |
| Publisher | nvidia |
|---|---|
| Origin | United States |
| Licence | Llama-3.1 |
| Publisher’s own licence statement | other |
| Parameters | 253,401,268,224 (253.4B) |
| Architecture | nemotron-nas |
| Layers | 162 |
| Hidden size | 16384 |
| Attention | MHA (128 heads, 128 KV) |
| Native precision | BF16 |
| Pulls recorded here | 36,006 |
| Downloads reported upstream | 1,634 |
| MMLU-Pro | 82.5 |
|---|---|
| GPQA Diamond | 72.8 |
| AIME 2025 | 63.7 |
| MATH-500 | 95.2 |
| LiveCodeBench | 64.1 |
| IFBench | 38.2 |
| Humanity's Last Exam | 7.4 |
Reported by a third party and recorded here, not re-run by us. Source: artificialanalysis.ai/api/v2.
Registry record as of 2026-09-01. Identity, licence and parameter count are properties of a release and do not move; tier can, and the verify link above re-checks it against the live registry rather than this page.