models · poolside

Laguna-S-2.1

Catalogued from local.ai local-inference benchmark index at Quarantine — unverified, pending attestation.

Compare Verify

What this registry has checked

Quarantine. We hold the publisher’s own published hash for these weights and have recorded it against this entry. We have not downloaded and hashed the weights ourselves, so the hash is the publisher’s claim, not our measurement.
Methodhf-published-sha256

What is recorded

Publisherpoolside
OriginNot recorded. The harvest wrote its own source name into this field, which tells us nothing about jurisdiction.
LicenceOpenMDW-1.1
Parameters117,561,977,600 (117.6B)
Context window262,144 tokens
Modalitytext->text
Architecturelaguna
Layers48
Hidden size3072
AttentionGQA (48 heads, 8 KV)
Native precisionBF16
Downloads reported upstream64,991

Measured on real hardware

None of these runs are ours. Each figure is the reporter’s, linked to their own evidence, under their own trust label. The number shown is single-stream decode — one user, concurrency 1 — because aggregate throughput across many users is a different quantity and reads as far faster than anyone will see.

Reported to local-ai-registry

HardwareEngineBuild measuredDecode, 1 streamLabelReported by
Radeon AI PRO R9700 ×3 devicesllama.cppQ4_0_ROCMFP4_COHERENT52.3 tok/s @ 2,048candidatelocalmaxxing
Radeon AI PRO R9700 ×3 devicesllama.cppQ4_0_ROCMFP4_STRIX_LEAN52.3 tok/s @ 2,048candidatelocalmaxxing
Radeon AI PRO R9700 ×3 devicesllama.cppQ4_0_ROCMFP4_FAST52.2 tok/s @ 2,048candidatelocalmaxxing
Measured outside 2K–32K context — decode speed depends heavily on context, so these are not comparable with the rows above
Radeon AI PRO R9700 ×3 devicesllama.cppUD-Q4_K_XL51.6 tok/s @ 131,072candidatelocalmaxxing
Ryzen AI Max+ 395llama.cppQ4_0_ROCMFP4_STRIX_LEAN41.7 tok/s @ 262,144candidatelocalmaxxing
Ryzen AI Max+ 395llama.cppQ4_0_ROCMFP4_FAST40.7 tok/s @ 820candidatelocalmaxxing
Ryzen AI Max+ 395llama.cppUD-Q4_K_XL29.9 tok/s @ 820candidatelocalmaxxing

From local-ai-registry (MIT), commit 124959d: 8 runs across 3 machines, 0 validated — model revision and runtime pinned, launch accepted — and 8 candidate, their label for useful evidence without a reproducible-launch promise.

Measured by local.ai

HardwareBuild measuredDecode @ 8KMemory @ 8KDecode mode
NVIDIA RTX PRO 6000 Blackwell 96GBQ4_K_M117 tok/s80.1 GBordinary
MacBook Pro M5 Max 128GB · 40-core GPUQ4_K_M48.3 tok/s87.0 GBordinary
Mac Studio M3 Ultra 96GB · 80-core GPUQ4_K_M47.4 tok/s88.7 GBordinary
Mac Studio M3 Ultra 96GB · 60-core GPUQ4_K_M47.1 tok/s78.1 GBordinary
NVIDIA DGX Spark 128GBQ4_K_M20.1 tok/s86.6 GBordinary
NVIDIA DGX Spark 128GBNVFP418.5 tok/s73.3 GBordinary

Measured by local.ai, not by us: 6 runs. Each names its engine pinned by image digest and the exact serve command, in the full record.

Elsewhere in the registry

Registry record as of 2026-09-26. Identity, licence and parameter count are properties of a release and do not move; tier can, and the verify link above re-checks it against the live registry rather than this page.