models · deepseek-ai
Qwen-32B distilled from DeepSeek-R1. Best open reasoning model at 32B scale.
| Weight file SHA-256 | 12075bd669f12c303058208404948ec47a626c75b347d4648b800f76a3b7b2c1 |
|---|---|
| Method | hf-published-sha256 |
| Publisher | deepseek-ai |
|---|---|
| Origin | China |
| Licence | MIT |
| Parameters | 32,763,876,352 (32.8B) |
| Architecture | qwen2 |
| Layers | 64 |
| Hidden size | 5120 |
| Attention | GQA (40 heads, 8 KV) |
| Native precision | BF16 |
| Pulls recorded here | 79,017 |
| Downloads reported upstream | 506,369 |
| MMLU-Pro | 73.9 |
|---|---|
| GPQA Diamond | 61.5 |
| AIME 2025 | 63 |
| MATH-500 | 94.1 |
| LiveCodeBench | 27 |
| IFBench | 22.9 |
| Humanity's Last Exam | 4.6 |
Reported by a third party and recorded here, not re-run by us. Source: artificialanalysis.ai/api/v2.
Registry record as of 2026-06-11. Identity, licence and parameter count are properties of a release and do not move; tier can, and the verify link above re-checks it against the live registry rather than this page.