models · deepseek-ai

DeepSeek-R1-Distill-Llama-70B

Llama-3-70B distilled from DeepSeek-R1 reasoning traces. Strong math and code at 70B scale.

Compare Verify

What this registry has checked

Quarantine. We hold the publisher’s own published hash for these weights and have recorded it against this entry. We have not downloaded and hashed the weights ourselves, so the hash is the publisher’s claim, not our measurement.
Weight file SHA-256e15541ba94dd226540a47bdec8ba43766e1144cc7007289b42712774363ff053
Methodhf-published-sha256

What is recorded

Publisherdeepseek-ai
OriginChina
LicenceMIT
Parameters70,553,706,496 (70.6B)
Context window8,192 tokens
Modalitytext->text
Architecturellama
Layers80
Hidden size8192
AttentionGQA (64 heads, 8 KV)
Native precisionBF16
Pulls recorded here84,017
Downloads reported upstream78,355

Reported scores

MMLU-Pro79.5
GPQA Diamond40.2
AIME 202553.7
MATH-50093.5
LiveCodeBench26.6
IFBench27.6
Humanity's Last Exam5.1

Reported by a third party and recorded here, not re-run by us. Source: artificialanalysis.ai/api/v2.

Elsewhere in the registry

Registry record as of 2026-06-11. Identity, licence and parameter count are properties of a release and do not move; tier can, and the verify link above re-checks it against the live registry rather than this page.