models · minimax

MiniMax M2.7

MiniMax M2.7. 204.8K context, 199 tok/s. Arena 1181. Latest generation MiniMax model with improved instruction following and reasoning.

Compare Verify

What this registry has checked

Quarantine. We hold the publisher’s own published hash for these weights and have recorded it against this entry. We have not downloaded and hashed the weights ourselves, so the hash is the publisher’s claim, not our measurement.
Weight file SHA-2567f54e26676152327d2cae77b05c8a642b70ee4c5a8ec49ede5cb254fbbd047d6
Methodhf-published-sha256

What is recorded

Publisherminimax
OriginRecorded as a covered nation under 10 U.S.C. § 4872(f) — the PRC, Russia, Iran or North Korea. The registry did not record which.
LicenceMiniMax-Open
Publisher’s own licence statementother
Parameters228,689,764,864 (228.7B)
Context window204.8K
Modalitytext->text
Architectureminimax_m2
Layers62
Hidden size3072
AttentionGQA (48 heads, 8 KV)
Native precisionF8_E4M3
Pulls recorded here270,015
Downloads reported upstream1,366,935

Reported scores

GPQA Diamond87.4
IFBench75.7
Humanity's Last Exam29.6

Reported by a third party and recorded here, not re-run by us. Source: artificialanalysis.ai/api/v2.

Measured on real hardware

None of these runs are ours. Each figure is the reporter’s, linked to their own evidence, under their own trust label. The number shown is single-stream decode — one user, concurrency 1 — because aggregate throughput across many users is a different quantity and reads as far faster than anyone will see.

Reported to local-ai-registry

HardwareEngineBuild measuredDecode, 1 streamLabelReported by
NVIDIA DGX Spark GB10llama.cppUD-IQ4_XS31.0 tok/s @ 32,768candidatelocalmaxxing
Intel Arc Pro B70 ×4 devicesllama.cppUD-IQ4_XS17.7 tok/s @ 2,048candidatelocalmaxxing
Ryzen AI Max+ 395llama.cppUD-Q3_K_M13.9 tok/s @ 16,384candidatelocalmaxxing
Measured outside 2K–32K context — decode speed depends heavily on context, so these are not comparable with the rows above
NVIDIA DGX Spark GB10llama.cppUD-IQ4_XS30.9 tok/s @ 108,000candidatelocalmaxxing
NVIDIA DGX Spark GB10vllmAWQ-4bit27.8 tok/s @ 190,000candidatelocalmaxxing
Intel Arc Pro B70 ×4 devicesllama.cppUD-IQ4_XS14.4 tok/s @ 64candidatelocalmaxxing

From local-ai-registry (MIT), commit 124959d: 21 runs across 6 machines, 0 validated — model revision and runtime pinned, launch accepted — and 21 candidate, their label for useful evidence without a reproducible-launch promise.

Elsewhere in the registry

Registry record as of 2026-08-13. Identity, licence and parameter count are properties of a release and do not move; tier can, and the verify link above re-checks it against the live registry rather than this page.