models · qwen

Qwen3.8-27B

image-text-to-text model harvested from Hugging Face (trending). Catalogued at Quarantine — unverified, pending attestation.

Compare Verify

What this registry has checked

Quarantine. We hold the publisher’s own published hash for these weights and have recorded it against this entry. We have not downloaded and hashed the weights ourselves, so the hash is the publisher’s claim, not our measurement.
Methodhf-published-sha256

What is recorded

Publisherqwen
OriginRecorded as a covered nation under 10 U.S.C. § 4872(f) — the PRC, Russia, Iran or North Korea. The registry did not record which.
Licenceapache-2.0
Parameters27,781,427,952 (27.8B)
Context window262,144 tokens
Modalitytext+image+video->text
Native precisionBF16
Pulls recorded here91,917
Downloads reported upstream7,703,400

Reported scores

GPQA Diamond90.5
Humanity's Last Exam33.9

Reported by a third party and recorded here, not re-run by us. Source: artificialanalysis.ai/api/v2.

Measured on real hardware

None of these runs are ours. Each figure is the reporter’s, linked to their own evidence, under their own trust label. The number shown is single-stream decode — one user, concurrency 1 — because aggregate throughput across many users is a different quantity and reads as far faster than anyone will see.

Reported to local-ai-registry

HardwareEngineBuild measuredDecode, 1 streamLabelReported by
GeForce RTX 3090 ×4 devicessglangAWQ INT4 W4A16137 tok/s @ 8,744validated0xsero
NVIDIA RTX PRO 4500 BlackwellTabbyAPI (ExLlamaV3)EXL3 SC 4.00bpw H5117 tok/s @ 32,768validated0xsero
GeForce RTX 5090llama.cppUD-Q4_K_XL116 tok/s @ 2,048candidatelocalmaxxing
GeForce RTX 5090llama.cppQ4_K_S102 tok/s @ 2,048candidatelocalmaxxing
GeForce RTX 5090llama.cppQ6_K97.7 tok/s @ 2,048candidatelocalmaxxing
Radeon RX 7900 XTXllama.cppQ4_K_XL90.7 tok/s @ 32,768candidatelocalmaxxing
Ryzen AI Max+ 395llama.cppIU490.1 tok/s @ 32,768candidatelocalmaxxing
GeForce RTX 5090ollamaQ4_K_M82.0 tok/s @ 16,384candidate0xsero
GeForce RTX 3080llama.cppIQ2_XXS74.1 tok/s @ 16,384candidatelocalmaxxing
GeForce RTX 3090 ×2 devicesbunn-llamaQ4_K_XL73.0 tok/s @ 2,048candidatelocalmaxxing

From local-ai-registry (MIT), commit 124959d: 251 runs across 45 machines, 49 validated — model revision and runtime pinned, launch accepted — and 202 candidate, their label for useful evidence without a reproducible-launch promise. 149 more not listed. 14 report figures that contradict each other — prefill below decode, which parallel prefill essentially never is — and are left to the source rather than ranked here.

Elsewhere in the registry

Registry record as of 2026-09-01. Identity, licence and parameter count are properties of a release and do not move; tier can, and the verify link above re-checks it against the live registry rather than this page.