April 23, 2024
Small and Mighty: NVIDIA Accelerates Microsoft’s Open Phi-3 Mini Language Models
NVIDIA announced today its acceleration of Microsoft’s new Phi-3 Mini open language model with NVIDIA TensorRT-LLM, an open-source library for optimizing large language model inference when running on NVIDIA GPUs from PC to cloud. Phi-3 Mini packs ...