models · deepseek-ai

DeepSeek-R1-Distill-Qwen-32B

Qwen-32B distilled from DeepSeek-R1. Best open reasoning model at 32B scale.

Compare Verify

What this registry has checked

Quarantine. We hold the publisher’s own published hash for these weights and have recorded it against this entry. We have not downloaded and hashed the weights ourselves, so the hash is the publisher’s claim, not our measurement.
Weight file SHA-25612075bd669f12c303058208404948ec47a626c75b347d4648b800f76a3b7b2c1
Methodhf-published-sha256

What is recorded

Publisherdeepseek-ai
OriginChina
LicenceMIT
Parameters32,763,876,352 (32.8B)
Architectureqwen2
Layers64
Hidden size5120
AttentionGQA (40 heads, 8 KV)
Native precisionBF16
Pulls recorded here79,017
Downloads reported upstream506,369

Reported scores

MMLU-Pro73.9
GPQA Diamond61.5
AIME 202563
MATH-50094.1
LiveCodeBench27
IFBench22.9
Humanity's Last Exam4.6

Reported by a third party and recorded here, not re-run by us. Source: artificialanalysis.ai/api/v2.

Elsewhere in the registry

Registry record as of 2026-06-11. Identity, licence and parameter count are properties of a release and do not move; tier can, and the verify link above re-checks it against the live registry rather than this page.