Amazon SageMaker AI Batch Transform Now Supports G6e Instances
Amazon SageMaker AI adds support for EC2 G6e instances in Batch Transform, enhancing performance for offline inference on large datasets
Amazon SageMaker AI now supports Amazon EC2 G6e instances for batch transforms, enabling offline inference for gpu-intensive workloads like large language models and diffusion models that generate images, video, or audio. The G6e instances feature up to eight NVIDIA L40S Tensor Core GPUs and third-generation AMD EPYC processors, making them ideal for large datasets that do not require a persistent inference endpoint. Supported regions include US East (N. Virginia), US East (Ohio), US West (Oregon), Asia Pacific (Mumbai), and Asia Pacific (Hyderabad).
Why it matters
This update affects users running large-scale offline inference tasks that require GPU resources. Users of Amazon SageMaker AI's batch transform functionality can now perform inference with higher performance.