datasets · huggingfacem4

The Cauldron

Aggregated multimodal vision-language instruction corpus spanning ~50 open VQA/captioning datasets, used to train Idefics2/3.

Verify

What this registry has checked

Quarantine. We hold the publisher’s own published hash for these weights and have recorded it against this entry. We have not downloaded and hashed the weights ourselves, so the hash is the publisher’s claim, not our measurement.
Weight file SHA-25667a0eb5299b76b79fec5a8512b16e3ad910342316522b87d80721b6b72699db9

What is recorded

Publisherhuggingfacem4
OriginEuropean Union
Licencevarious-documented
Modalityimage-text
Pulls recorded here19,300

Elsewhere in the registry

Registry record as of 2026-06-06. Identity, licence and parameter count are properties of a release and do not move; tier can, and the verify link above re-checks it against the live registry rather than this page.