systems · nvidia

TensorRT-LLM

Compiled inference engine with kernel fusion, in-flight batching, and FP8/FP4 paths. Top throughput on NVIDIA GPUs.

What this registry has checked

Quarantine. We hold the publisher’s own published hash for these weights and have recorded it against this entry. We have not downloaded and hashed the weights ourselves, so the hash is the publisher’s claim, not our measurement.
Weight file SHA-256754595d3bef49bf4fcdf6e047dfae9540ce81370d37b67c5f8f947fcea559437

What is recorded

Publishernvidia
OriginUnited States
LicenceApache-2.0
Pulls recorded here320,002

Elsewhere in the registry

Registry record as of 2026-06-06. Identity, licence and parameter count are properties of a release and do not move; tier can, and the verify link above re-checks it against the live registry rather than this page.