models · google
Google Gemma 3 compact 4B instruct model. Multimodal (text + image). Designed for edge and low-resource deployments.
| Weight file SHA-256 | 125a8f8f23ceea7ab698640c72cd6de9fe05cb4e0983d185987c2c3c628ed1b3 |
|---|---|
| Method | hf-published-sha256 |
| Publisher | |
|---|---|
| Origin | United States |
| Licence | Gemma |
| Parameters | 4,300,079,472 (4.3B) |
| Context window | 128K |
| Modality | text+image->text |
| Native precision | BF16 |
| Pulls recorded here | 720,019 |
| Downloads reported upstream | 1,709,353 |
| Hardware | Engine | Build measured | Decode, 1 stream | Label | Reported by |
|---|---|---|---|---|---|
| Apple M3 Ultra 512GB | llama.cpp | Q4_K_M | 127 tok/s @ 8,192 | candidate | localmaxxing |
| GeForce RTX 3060 | llama.cpp | Q4_K_M | 97.0 tok/s @ 4,608 | candidate | localmaxxing |
From local-ai-registry (MIT), commit 124959d: 2 runs across 2 machines, 0 validated — model revision and runtime pinned, launch accepted — and 2 candidate, their label for useful evidence without a reproducible-launch promise.
Registry record as of 2026-07-25. Identity, licence and parameter count are properties of a release and do not move; tier can, and the verify link above re-checks it against the live registry rather than this page.