NVIDIA
NVIDIA
Phi-2 (TensorRT LLM)
Model
NVIDIA
NVIDIA
Phi-2 (TensorRT LLM)

Phi-2 is a 2.7 billion parameter language model developed by Microsoft Research. The phi-2 model is best suited for prompts using the Question-Answer (QA) format, the chat format, and the code format.

6 Versions
06/13/2024 11:30 PM UTC1.47 GB
TRT Cloud Engine Properties
KeyValue
cpu_archx86_64
osWindows
num_gpus1
trtllm_version0.10.0
max_batch_size1
weightlessFalse
gpuRTX4090
06/13/2024 11:29 PM UTC1.47 GB
TRT Cloud Engine Properties
KeyValue
cpu_archx86_64
osWindows
num_gpus1
trtllm_version0.10.0
weight_strippedFalse
max_batch_size1
gpuRTX4070
06/13/2024 11:29 PM UTC1.47 GB
TRT Cloud Engine Properties
KeyValue
cpu_archx86_64
osWindows
num_gpus1
trtllm_version0.10.0
weight_strippedFalse
max_batch_size1
gpuRTX4060TI
06/13/2024 11:28 PM UTC1.47 GB
TRT Cloud Engine Properties
KeyValue
cpu_archx86_64
osWindows
num_gpus1
trtllm_version0.10.0
max_batch_size1
weightlessFalse
gpuRTX3070
06/13/2024 11:28 PM UTC4.78 GB
TRT Cloud Engine Properties
KeyValue
cpu_archx86_64
osWindows
num_gpus1
trtllm_version0.10.0
max_batch_size1
weightlessFalse
gpuRTX4090
06/13/2024 11:27 PM UTC4.78 GB
TRT Cloud Engine Properties
KeyValue
cpu_archx86_64
osWindows
num_gpus1
trtllm_version0.10.0
weight_strippedFalse
max_batch_size1
gpuRTX4070