Skip to main content
NVIDIA
TensorRT LLM Develop
Container
NVIDIA
TensorRT LLM Develop

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs.

  • LayerLabelCreated
    sha256:3f9cda79bd817d1b6cc7e9fad97c0e7c5e4ce7a39c17d5ca2e7a8b9d4a31f2bcRUN
    GITHUB_MIRROR=https://urm.nvidia.com/artifactory/github-go-remote PYTHON_VERSION=3.12.3 TRT_VER= CUDA_VER= CUDNN_VER= NCCL_VER= CUBLAS_VER= TORCH_INSTALL_TYPE=skip /bin/bash -c pip3 install --upgrade --no-cache-dir "protobuf>=4.25.8"
    09/11/2025 2:12 PM UTC
    sha256:6cc08a4730ba887ca73c56df4e7acac1e3873e33c85f809d75e34d24f3a0e20aRUN
    GITHUB_MIRROR=https://urm.nvidia.com/artifactory/github-go-remote PYTHON_VERSION=3.12.3 TRT_VER= CUDA_VER= CUDNN_VER= NCCL_VER= CUBLAS_VER= TORCH_INSTALL_TYPE=skip /bin/bash -c pip3 uninstall -y opencv &&
      rm -rf /usr/local/lib/python3*/dist-packages/cv2/ &&
      pip3 install opencv-python-headless --force-reinstall --no-deps --no-cache-dir
    09/11/2025 2:10 PM UTC
    sha256:a3ed95caeb02ffe68cdd9fd84406680ae93d633cb16422d00e8a7c22955b46d4ENV
    PYTORCH_CUDA_ALLOC_CONF=garbage_collection_threshold:0.99999
    09/11/2025 1:39 PM UTC
    sha256:39db0bb24db3080cbfc33ec2791f551bdd5626e41568fbd102352a01be8e164eRUN
    GITHUB_MIRROR=https://urm.nvidia.com/artifactory/github-go-remote PYTHON_VERSION=3.12.3 TRT_VER= CUDA_VER= CUDNN_VER= NCCL_VER= CUBLAS_VER= TORCH_INSTALL_TYPE=skip /bin/bash -c bash ./install_pytorch.sh $TORCH_INSTALL_TYPE &&
      rm install_pytorch.sh
    09/11/2025 1:39 PM UTC
    sha256:8d019f9ad06f538fdb38ea963219797ebf09e42f240d715e00d305cbdf61ae96COPY
    docker/common/install_pytorch.sh install_pytorch.sh
    09/11/2025 1:39 PM UTC
    sha256:a3ed95caeb02ffe68cdd9fd84406680ae93d633cb16422d00e8a7c22955b46d4ARG
    TORCH_INSTALL_TYPE=skip
    09/11/2025 1:39 PM UTC
    sha256:782116fbd2861a8fcbd62c4b345a1a15f6384d92eea0dbb5d0cd370aa3f1ca7eRUN
    GITHUB_MIRROR=https://urm.nvidia.com/artifactory/github-go-remote PYTHON_VERSION=3.12.3 TRT_VER= CUDA_VER= CUDNN_VER= NCCL_VER= CUBLAS_VER= /bin/bash -c GITHUB_MIRROR=$GITHUB_MIRROR bash ./install_mpi4py.sh &&
      rm install_mpi4py.sh
    09/11/2025 1:39 PM UTC
    sha256:34ba2726bed6ae3911c8f25b64fc2a5ecab795f5c3bc0f39a3ddaa4df832525eCOPY
    docker/common/install_mpi4py.sh install_mpi4py.sh
    09/11/2025 1:20 PM UTC
    sha256:e7e4742184182ec6a8ec670473f08426f886a95e0c2f14a1993c8edfa004d6acRUN
    GITHUB_MIRROR=https://urm.nvidia.com/artifactory/github-go-remote PYTHON_VERSION=3.12.3 TRT_VER= CUDA_VER= CUDNN_VER= NCCL_VER= CUBLAS_VER= /bin/bash -c bash ./install_polygraphy.sh &&
      rm install_polygraphy.sh
    09/11/2025 1:20 PM UTC
    sha256:5ddfba05cd8c5b243bd7fa79f1b96493856e53813035f1b1886793ecf58808b8COPY
    docker/common/install_polygraphy.sh install_polygraphy.sh
    09/11/2025 1:17 PM UTC
    ...

    NVIDIA uses cookies to improve your experience on our web site. We and our third-party partners also use cookies and other tools to collect and record information you provide as well as information about your interactions with our websites for performance improvement, analytics, and to assist in marketing efforts. By clicking "Accept All", you consent to our use of cookies and other tools as described in our Cookie Policy. You can manage your cookie settings by clicking on "Manage Settings." By continuing to use this site or by clicking one of the buttons below, you agree to our Terms of Service (which contains important waivers). Please see our Privacy Policy for more information on our privacy practices.