models · qwen
Dense 72B instruction-tuned model with strong multilingual, coding, and long-context performance.
| Weight file SHA-256 | 0216de5a9bb088c70f5ed71bae363f22517bce296e399d3eff441ac94b79144b |
|---|---|
| Method | hf-published-sha256 |
| Publisher | qwen |
|---|---|
| Origin | Recorded as a covered nation under 10 U.S.C. § 4872(f) — the PRC, Russia, Iran or North Korea. The registry did not record which. |
| Licence | Qwen |
| Publisher’s own licence statement | other |
| Parameters | 72,706,203,648 (72.7B) |
| Context window | 128K |
| Modality | text->text |
| Architecture | qwen2 |
| Layers | 80 |
| Hidden size | 8192 |
| Attention | GQA (64 heads, 8 KV) |
| Native precision | BF16 |
| Pulls recorded here | 760,016 |
| Downloads reported upstream | 417,294 |
| MMLU-Pro | 72 |
|---|---|
| GPQA Diamond | 49.1 |
| AIME 2025 | 14 |
| MATH-500 | 85.8 |
| LiveCodeBench | 27.6 |
| IFBench | 36.9 |
| Humanity's Last Exam | 3.6 |
Reported by a third party and recorded here, not re-run by us. Source: artificialanalysis.ai/api/v2.
Registry record as of 2026-06-11. Identity, licence and parameter count are properties of a release and do not move; tier can, and the verify link above re-checks it against the live registry rather than this page.