models · qwen
Qwen2.5 math-specialized 72B instruct model. Optimized for mathematical reasoning, AIME, and MATH-500 benchmarks.
| Weight file SHA-256 | 21fad6da9da2513ce8bf66b9e2256bd1e99edb69393a9c6ec4cf9b98ab0d8df5 |
|---|---|
| Method | hf-published-sha256 |
| Publisher | qwen |
|---|---|
| Origin | Recorded as a covered nation under 10 U.S.C. § 4872(f) — the PRC, Russia, Iran or North Korea. The registry did not record which. |
| Licence | Apache-2.0 |
| Publisher’s own licence statement | other |
| Parameters | 72,706,203,648 (72.7B) |
| Context window | 4K |
| Architecture | qwen2 |
| Layers | 80 |
| Hidden size | 8192 |
| Attention | GQA (64 heads, 8 KV) |
| Native precision | BF16 |
| Pulls recorded here | 310,024 |
| Downloads reported upstream | 836 |
Registry record as of 2026-06-11. Identity, licence and parameter count are properties of a release and do not move; tier can, and the verify link above re-checks it against the live registry rather than this page.