architectures · minimax
MiniMax hybrid architecture combining linear Lightning Attention with standard softmax attention in alternating layers. Enables O(1) per-token inference cost at 1M context without KV-cache growth. Used in MiniMax M2 through M2.7. Achieves 400+ tok/s throughput at 230B scale.
| Weight file SHA-256 | c321fb89f473be5e518d8ca15faa86a0f6b39797977b0ac55ac28ad9dfb7bae5 |
|---|
| Publisher | minimax |
|---|---|
| Origin | Recorded as a covered nation under 10 U.S.C. § 4872(f) — the PRC, Russia, Iran or North Korea. The registry did not record which. |
| Licence | MiniMax-Open |
| Pulls recorded here | 51,918 |
Registry record as of 2026-08-14. Identity, licence and parameter count are properties of a release and do not move; tier can, and the verify link above re-checks it against the live registry rather than this page.