21:37
2026-09-10
aws.amazon.com
ai-infrastructure
Reduce inference cold starts on Amazon SageMaker HyperPod with model caching
Amazon launched model caching for Amazon SageMaker Inference on HyperPod, a feature that pre-loads model weights and container images onto cluster nodes so inference pods can read from local NVMe storβ¦