datasets · huggingfacem4
Aggregated multimodal vision-language instruction corpus spanning ~50 open VQA/captioning datasets, used to train Idefics2/3.
| Weight file SHA-256 | 67a0eb5299b76b79fec5a8512b16e3ad910342316522b87d80721b6b72699db9 |
|---|
| Publisher | huggingfacem4 |
|---|---|
| Origin | European Union |
| Licence | various-documented |
| Modality | image-text |
| Pulls recorded here | 19,300 |
Registry record as of 2026-06-06. Identity, licence and parameter count are properties of a release and do not move; tier can, and the verify link above re-checks it against the live registry rather than this page.