USASI
SoftwareRuntime

TensorRT-LLM

Project record

Maintained by NVIDIA12

TensorRT-LLM is NVIDIA's open-source library for running large language model inference on NVIDIA GPUs. It provides a Python LLM API, optimized kernels, Python and C++ runtimes, and an online serving command (trtllm-serve).24

Last reviewedEntry updated Documented release Unknown

Availability and license

Overall availability

Public

Source code is public on GitHub; prebuilt wheels are installable with pip and development containers are available from NVIDIA NGC.25

Availability is separate from permission: read the license before using or redistributing.

The license file states the project is under Apache 2.0 but contains portions derived from other open-source projects that may carry different licenses, listed in the same file; individual file headers give specific terms.1

Public materials checklist

Items for a runtime under USASI rubric v0.1. Unknown means unassessed or insufficient evidence.
Public materials checklist for TensorRT-LLM
ItemStatusNotes and evidence
Source codeIs the source code publicly readable?Public2
DocumentationIs user documentation published?Public4
InstallationAre installation instructions or packages publicly available?PublicInstallation via a PyPI wheel, NGC container images, or building from source on Linux.5
Supported platformsAre supported operating systems or hardware documented?PublicThe support matrix lists Linux x86_64 and aarch64 and NVIDIA Ampere through Blackwell GPU architectures.6
Release statusAre versioned releases published?PublicVersioned releases are published on GitHub; v1.2.1 (April 2026) was the latest non-prerelease when checked.3

What it is useful for

The documentation covers offline inference through the LLM API, online serving, quantization, speculative decoding, KV cache management, LoRA adapters, and parallelism strategies for deploying models on NVIDIA GPUs.4

Organization context

U.S. eligibility

Project eligibility rests on documented governing or maintaining entities, not on contributors.

Eligible · basis: U.S.-governed project

The repository is published under the NVIDIA GitHub organization and its license file names NVIDIA Corporation & Affiliates as copyright holder. NVIDIA's principal executive offices are in Santa Clara, California, per its Form 10-Q for the quarter ended July 26, 2026.127

Assessed Sep 29, 2026

Sources

  1. 1.
    TensorRT-LLM LICENSE (external site: raw.githubusercontent.com)

    NVIDIA (GitHub) · License · accessed Sep 29, 2026

  2. 2.
  3. 3.
    GitHub releases for NVIDIA/TensorRT-LLM (external site: api.github.com)

    GitHub · Release notes · accessed Sep 29, 2026

  4. 4.
    TensorRT LLM documentation (external site: nvidia.github.io)

    NVIDIA · Documentation · accessed Sep 29, 2026

  5. 5.
  6. 6.
    Support Matrix - TensorRT LLM (external site: nvidia.github.io)

    NVIDIA · Documentation · accessed Sep 29, 2026

  7. 7.
    NVIDIA Corporation Form 10-Q for the quarter ended July 26, 2026 (external site: sec.gov)

    U.S. Securities and Exchange Commission (EDGAR) · Filing · published Aug 26, 2026 · accessed Sep 29, 2026

This listing is not an endorsement, a safety assessment, or a federal approval.

Support Us

Help keep USASI useful.

Optional. No USASI account required. Payment takes place on the linked provider’s website (Buy Me a Coffee).

About supporting this project