models · qwen

Qwen3-4B

Compact 4B Qwen3 with hybrid thinking. Outperforms much larger models on math benchmarks.

Compare Verify

What this registry has checked

Quarantine. We hold the publisher’s own published hash for these weights and have recorded it against this entry. We have not downloaded and hashed the weights ourselves, so the hash is the publisher’s claim, not our measurement.
Weight file SHA-2566912e3a1fafaa368b0b53a0c13fee4803fd0c8017049157d15a89cbe996af020
Methodhf-published-sha256

What is recorded

Publisherqwen
OriginChina
LicenceApache-2.0
Parameters4,022,468,096 (4.02B)
Architectureqwen3
Layers36
Hidden size2560
AttentionGQA (32 heads, 8 KV)
Native precisionBF16
Pulls recorded here67,020
Downloads reported upstream6,204,889

Reported scores

MMLU-Pro58.6
GPQA Diamond39.8
AIME 202558.3
MATH-50084.3
LiveCodeBench23.3
Humanity's Last Exam3.3

Reported by a third party and recorded here, not re-run by us. Source: artificialanalysis.ai/api/v2.

Elsewhere in the registry

Registry record as of 2026-09-15. Identity, licence and parameter count are properties of a release and do not move; tier can, and the verify link above re-checks it against the live registry rather than this page.