Announcing region expansion of G6 instances on SageMaker AI Inference

Amazon EC2 G6 Instances Now Available in AWS GovCloud (US-East) on Amazon SageMaker AI Inference
Amazon EC2 G6 instances, powered by up to 8 NVIDIA L4 Tensor Core GPUs and third-generation AMD EPYC processors, are now available in the AWS GovCloud (US-East) region on Amazon SageMaker AI inference. These instances deliver up to 2x the deep learning inference performance compared to G4dn instances.
This region expansion enables government agencies and organizations in GovCloud to deploy inference endpoints on G6 instances for generative AI workloads, including small-to-medium language models, image generation, and computer vision tasks, while meeting compliance and data residency requirements.
G6 instances offer strong price-performance for production inference workloads that fit within 24 GB of GPU memory.
What to do
- Deploy inference endpoints on G6 instances for generative AI workloads.
- Visit the pricing page for more information.
Source: AWS release notes
If you need further guidance on AWS, our experts are available at AWS@westloop.io. You may also reach us by submitting the Contact Us form.



