Announcing region expansion of G6 instances on SageMaker AI Inference
Share
Services
We are pleased to announce the availability of [Amazon EC2 G6 ](https://aws.amazon.com/ec2/instance-types/g6/)instances in the AWS GovCloud (US-East) region on Amazon SageMaker AI inference. G6 instances are powered by up to 8 NVIDIA L4 Tensor Core GPUs, each with 24 GB of memory, and third-generation AMD EPYC processors, delivering up to 2x the deep learning inference performance compared to G4dn instances.
With this region expansion, government agencies and organizations operating in GovCloud can deploy inference endpoints on G6 instances to serve generative AI workloads—including small-to-medium language models, image generation, and computer vision tasks—while meeting strict compliance and data residency requirements. G6 instances offer strong price-performance for production inference workloads that fit within 24 GB of GPU memory.
G6 instances for SageMaker AI inference are now available in AWS GovCloud (US-East), in addition to previously supported regions. For pricing information on these instances, please visit our [pricing page](https://aws.amazon.com/sagemaker/ai/pricing/?refid=ft%5Fsagemaker).
What else is happening at Amazon Web Services?
Amazon Bedrock AgentCore now delivers unified observability with traces and logs in a single log group
about 4 hours ago
Services
Share
Read update
Services
Share
Read update
Services
Share
Read update
Services
Share
Read update
Services
Share
Read update
Services
Share