Announcing region expansion of G6 instances on SageMaker AI Inference

Published
July 23, 2026
https://aws.amazon.com/about-aws/whats-new/2026/07/g6-sagemaker-ai-inference/

Amazon EC2 G6 Instances Now Available in AWS GovCloud (US-East) on Amazon SageMaker AI Inference

Amazon EC2 G6 instances, powered by up to 8 NVIDIA L4 Tensor Core GPUs and third-generation AMD EPYC processors, are now available in the AWS GovCloud (US-East) region on Amazon SageMaker AI inference. These instances deliver up to 2x the deep learning inference performance compared to G4dn instances.

This region expansion enables government agencies and organizations in GovCloud to deploy inference endpoints on G6 instances for generative AI workloads, including small-to-medium language models, image generation, and computer vision tasks, while meeting compliance and data residency requirements.

G6 instances offer strong price-performance for production inference workloads that fit within 24 GB of GPU memory.

What to do

  • Deploy inference endpoints on G6 instances for generative AI workloads.
  • Visit the pricing page for more information.

Source: AWS release notes




If you need further guidance on AWS, our experts are available at AWS@westloop.io. You may also reach us by submitting the Contact Us form.

Follow our blog

Get the latest insights and advice on AWS services from our experts.

By clicking Sign Up you're confirming that you agree with our Terms and Conditions.
Thank you! Your submission has been received!
Oops! Something went wrong while submitting the form.