NVIDIA
TensorRT LLM Develop
Container
NVIDIA
TensorRT LLM Develop

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs.

  • LayerLabelCreated
    sha256:0a9c183820e7371b214406c02db1309b2c3d8521c9d6f11d4e5e90954456c4f2RUN
    GITHUB_MIRROR=https://urm.nvidia.com/artifactory/github-go-remote PYTHON_VERSION=3.12.3 TRT_VER= CUDA_VER= CUDNN_VER= NCCL_VER= CUBLAS_VER= TORCH_INSTALL_TYPE=skip /bin/bash -c pip3 install --no-cache-dir -r /tmp/constraints.txt &&
      rm /tmp/constraints.txt
    12/19/2025 2:33 AM UTC
    sha256:bcefe12b1d566e6507a2fd7e0ed95ceddba7ab0e2009a723afa0f7f799118dabCOPY
    constraints.txt /tmp/constraints.txt
    12/19/2025 2:31 AM UTC
    sha256:6df4131d2483f279b9059246c0aa9f8f02e8c9c85dd0cb22ff30c0bf8b3dd32cRUN
    GITHUB_MIRROR=https://urm.nvidia.com/artifactory/github-go-remote PYTHON_VERSION=3.12.3 TRT_VER= CUDA_VER= CUDNN_VER= NCCL_VER= CUBLAS_VER= TORCH_INSTALL_TYPE=skip /bin/bash -c bash ./install.sh --opencv &&
      rm install.sh
    12/19/2025 2:31 AM UTC
    sha256:55ab854ec9ce98f234367d30ef5dfb88458a76a4c333f463dafa7824f435a51eRUN
    GITHUB_MIRROR=https://urm.nvidia.com/artifactory/github-go-remote PYTHON_VERSION=3.12.3 TRT_VER= CUDA_VER= CUDNN_VER= NCCL_VER= CUBLAS_VER= TORCH_INSTALL_TYPE=skip /bin/bash -c TORCH_INSTALL_TYPE=${TORCH_INSTALL_TYPE} bash ./install.sh --pytorch &&
      rm install_pytorch.sh
    12/19/2025 2:29 AM UTC
    sha256:a3ed95caeb02ffe68cdd9fd84406680ae93d633cb16422d00e8a7c22955b46d4ARG
    TORCH_INSTALL_TYPE=skip
    12/19/2025 2:29 AM UTC
    sha256:83a3aae1fd95f2fe423f1ab05639162c59198fe0e3646b66d228ad633f66727eRUN
    GITHUB_MIRROR=https://urm.nvidia.com/artifactory/github-go-remote PYTHON_VERSION=3.12.3 TRT_VER= CUDA_VER= CUDNN_VER= NCCL_VER= CUBLAS_VER= /bin/bash -c GITHUB_MIRROR=${GITHUB_MIRROR} bash ./install.sh --mpi4py &&
      rm install_mpi4py.sh
    12/19/2025 2:29 AM UTC
    sha256:6cb6adbb36d1b03f281c7976facb070eb0a65386093c63437fb39c043f0af37bRUN
    GITHUB_MIRROR=https://urm.nvidia.com/artifactory/github-go-remote PYTHON_VERSION=3.12.3 TRT_VER= CUDA_VER= CUDNN_VER= NCCL_VER= CUBLAS_VER= /bin/bash -c bash ./install.sh --polygraphy &&
      rm install_polygraphy.sh
    12/19/2025 2:24 AM UTC
    sha256:36b27e42e210aba705d63f39275cf9fa8962e46b74a53978a3f92a4e98cf5cffRUN
    GITHUB_MIRROR=https://urm.nvidia.com/artifactory/github-go-remote PYTHON_VERSION=3.12.3 TRT_VER= CUDA_VER= CUDNN_VER= NCCL_VER= CUBLAS_VER= /bin/bash -c TRT_VER=${TRT_VER} CUDA_VER=${CUDA_VER} CUDNN_VER=${CUDNN_VER} NCCL_VER=${NCCL_VER} CUBLAS_VER=${CUBLAS_VER} bash ./install.sh --tensorrt &&
      rm install_tensorrt.sh
    12/19/2025 2:24 AM UTC
    sha256:a3ed95caeb02ffe68cdd9fd84406680ae93d633cb16422d00e8a7c22955b46d4ARG
    CUBLAS_VER
    12/19/2025 2:21 AM UTC
    sha256:a3ed95caeb02ffe68cdd9fd84406680ae93d633cb16422d00e8a7c22955b46d4ARG
    NCCL_VER
    12/19/2025 2:21 AM UTC
    ...

    NVIDIA uses cookies to improve your experience on our web site. We and our third-party partners also use cookies and other tools to collect and record information you provide as well as information about your interactions with our websites for performance improvement, analytics, and to assist in marketing efforts. By clicking "Accept All", you consent to our use of cookies and other tools as described in our Cookie Policy. You can manage your cookie settings by clicking on "Manage Settings." By continuing to use this site or by clicking one of the buttons below, you agree to our Terms of Service (which contains important waivers). Please see our Privacy Policy for more information on our privacy practices.