21:51
2026-09-21
developer.nvidia.com
ai-infrastructure
Simplifying Model Serving Across Multiple GPUs with NVIDIA TensorRT Multi-Device Integration in NVIDIA Dynamo-Triton
NVIDIA Dynamo-Triton release 26.07 now enables TensorRT multi-device inference, allowing a single TensorRT network to execute across multiple GPUs using NCCL-backed distributed collectives, with full …