models · deepseek-ai
Catalogued from llm-stats Open LLM Leaderboard at Quarantine — unverified, pending attestation.
| Method | hf-published-sha256 |
|---|
| Publisher | deepseek-ai |
|---|---|
| Origin | Recorded as a covered nation under 10 U.S.C. § 4872(f) — the PRC, Russia, Iran or North Korea. The registry did not record which. |
| Licence | MIT |
| Parameters | 684,531,386,000 (684.5B) |
| Context window | 163,840 tokens |
| Modality | text->text |
| Architecture | deepseek_v3 |
| Layers | 61 |
| Hidden size | 7168 |
| Attention | MLA (128 heads, 128 KV) |
| Native precision | F8_E4M3 |
| Pulls recorded here | 6 |
| Downloads reported upstream | 249,493 |
| MMLU-Pro | 83.3 |
|---|---|
| GPQA Diamond | 73.5 |
| AIME 2025 | 49.7 |
| LiveCodeBench | 57.7 |
| IFBench | 37.8 |
| Humanity's Last Exam | 6.7 |
Reported by a third party and recorded here, not re-run by us. Source: artificialanalysis.ai/api/v2.
Registry record as of 2026-09-26. Identity, licence and parameter count are properties of a release and do not move; tier can, and the verify link above re-checks it against the live registry rather than this page.