optimizers · google

LAMB

Layer-wise Adaptive Moments for large-batch training (You et al., 2019). Enabled BERT pretraining in 76 minutes at batch size 32k.

What this registry has checked

Quarantine. We hold the publisher’s own published hash for these weights and have recorded it against this entry. We have not downloaded and hashed the weights ourselves, so the hash is the publisher’s claim, not our measurement.
Weight file SHA-256b1b57872d4d3d9f967dae7e540ab5c3eeae2e235a826b0eb160352bbdf753942

What is recorded

Publishergoogle
OriginUnited States
LicenceApache-2.0
Pulls recorded here44,302

Elsewhere in the registry

Registry record as of 2026-06-06. Identity, licence and parameter count are properties of a release and do not move; tier can, and the verify link above re-checks it against the live registry rather than this page.