SageMaker AI Inference Now Available with G6 Instances in AWS GovCloud (US-East)
Amazon SageMaker introduces G6 instances in AWS GovCloud (US-East), enabling government agencies and organizations to run generative AI workloads while meeting strict compliance and data residency requirements
Amazon EC2 G6 instances are now available for Amazon SageMaker AI inference in the AWS GovCloud (US-East) region. G6 instances feature up to 8 NVIDIA L4 Tensor Core GPUs, each with 24 GB of memory, and third-generation AMD EPYC processors, delivering up to 2x the deep learning inference performance compared to G4dn instances. This region expansion allows government agencies and organizations operating in GovCloud to deploy inference endpoints on G6 instances for generative AI workloads-including small-to-medium language models, image generation, and computer vision tasks-while meeting strict compliance and data residency requirements. G6 instances offer strong price-performance for production inference workloads that fit within 24 GB of GPU memory.