architectures · google
The original encoder-decoder self-attention architecture (Vaswani et al., 2017). Foundation of nearly all modern sequence models.
| Weight file SHA-256 | a7468c68516526913141989e07fcccd64bf51644bd7287e30a432b7a17265e59 |
|---|
| Publisher | |
|---|---|
| Origin | United States |
| Licence | Apache-2.0 |
| Pulls recorded here | 320,002 |
Registry record as of 2026-06-06. Identity, licence and parameter count are properties of a release and do not move; tier can, and the verify link above re-checks it against the live registry rather than this page.