models · poolside
Catalogued from local.ai local-inference benchmark index at Quarantine — unverified, pending attestation.
| Method | hf-published-sha256 |
|---|
| Publisher | poolside |
|---|---|
| Origin | Not recorded. The harvest wrote its own source name into this field, which tells us nothing about jurisdiction. |
| Licence | OpenMDW-1.1 |
| Parameters | 33,442,617,088 (33.4B) |
| Context window | 262,144 tokens |
| Modality | text->text |
| Architecture | laguna |
| Layers | 40 |
| Hidden size | 2048 |
| Attention | GQA (48 heads, 8 KV) |
| Native precision | BF16 |
| Downloads reported upstream | 63,504 |
| Hardware | Engine | Build measured | Decode, 1 stream | Label | Reported by |
|---|---|---|---|---|---|
| Intel Arc Pro B70 | llama.cpp | Q4_K_M | 87.8 tok/s @ 2,048 | candidate | localmaxxing |
| Intel Arc Pro B60 | llama.cpp | Q4_K_M | 72.5 tok/s @ 2,048 | candidate | localmaxxing |
| Measured outside 2K–32K context — decode speed depends heavily on context, so these are not comparable with the rows above | |||||
| Radeon AI PRO R9700 ×3 devices | llama.cpp | Q8_0 | 100 tok/s @ 792 | candidate | localmaxxing |
From local-ai-registry (MIT), commit 124959d: 4 runs across 4 machines, 0 validated — model revision and runtime pinned, launch accepted — and 4 candidate, their label for useful evidence without a reproducible-launch promise.
| Hardware | Build measured | Decode @ 8K | Memory @ 8K | Decode mode |
|---|---|---|---|---|
| NVIDIA RTX PRO 6000 Blackwell 96GB | Q4_K_M | 263 tok/s | 32.0 GB | ordinary |
| MacBook Pro M5 Max 128GB · 40-core GPU | Q4_K_M | 145 tok/s | 27.2 GB | ordinary |
| Mac Studio M3 Ultra 96GB · 80-core GPU | Q4_K_M | 107 tok/s | 31.6 GB | ordinary |
| MacBook Pro M4 Max 128GB · 40-core GPU | Q4_K_M | 106 tok/s | 29.4 GB | ordinary |
| Mac Studio M3 Ultra 96GB · 80-core GPU | Q4_K_M | 102 tok/s | 31.6 GB | ordinary |
| Mac Studio M3 Ultra 96GB · 60-core GPU | Q4_K_M | 101 tok/s | 31.2 GB | ordinary |
| NVIDIA DGX Spark 128GB | Q4_K_M | 79.5 tok/s | 31.6 GB | ordinary |
| MacBook Pro M5 Pro 64GB · 20-core GPU | Q4_K_M | 78.2 tok/s | 29.6 GB | ordinary |
Measured by local.ai, not by us: 15 runs, 7 more not listed. Each names its engine pinned by image digest and the exact serve command, in the full record.
Registry record as of 2026-09-26. Identity, licence and parameter count are properties of a release and do not move; tier can, and the verify link above re-checks it against the live registry rather than this page.