models · allenai
Olmo 3 32B Think — catalogued from third-party evaluation data; no weights held by Sovereign Frontier.
| Method | hf-published-sha256 |
|---|
| Publisher | allenai |
|---|---|
| Origin | United States |
| Licence | Apache-2.0 |
| Parameters | 32,233,522,176 (32.2B) |
| Context window | 65,536 tokens |
| Modality | text->text |
| Architecture | olmo3 |
| Layers | 64 |
| Hidden size | 5120 |
| Attention | GQA (40 heads, 8 KV) |
| Native precision | BF16 |
| Pulls recorded here | 6 |
| Downloads reported upstream | 9,502 |
| MMLU-Pro | 75.9 |
|---|---|
| GPQA Diamond | 61 |
| AIME 2025 | 73.7 |
| LiveCodeBench | 67.2 |
| IFBench | 49.1 |
| Humanity's Last Exam | 6.4 |
Reported by a third party and recorded here, not re-run by us. Source: artificialanalysis.ai/api/v2.
| Hardware | Engine | Build measured | Decode, 1 stream | Label | Reported by |
|---|---|---|---|---|---|
| GeForce RTX 3090 | llama.cpp | Q4_K_M | 23.1 tok/s @ 512 | candidate | localmaxxing |
| GeForce RTX 3060 ×2 devices | llama.cpp | Q4_K_M | 16.4 tok/s @ 512 | candidate | localmaxxing |
From local-ai-registry (MIT), commit 124959d: 2 runs across 2 machines, 0 validated — model revision and runtime pinned, launch accepted — and 2 candidate, their label for useful evidence without a reproducible-launch promise.
Registry record as of 2026-08-13. Identity, licence and parameter count are properties of a release and do not move; tier can, and the verify link above re-checks it against the live registry rather than this page.