Amazon SageMaker AI Batch Transform now supports G6e instances

Amazon SageMaker AI now supports Amazon EC2 G6e instances for batch transform. Batch Transform enables you to run predictions on datasets stored in Amazon S3 and is suited for large datasets that do not require a persistent inference endpoint.

Amazon EC2 G6e instances are powered by up to eight NVIDIA L40S Tensor Core GPUs with 48 GB of memory per GPU and third-generation AMD EPYC processors. G6e instances deliver improved performance for GPU-intensive workloads. With this launch, you can use G6e instances for GPU-intensive offline inference workloads, including large language models and diffusion models that generate images, video, and audio. To get started, select a supported ml.g6e instance type when creating a Batch Transform job through AWS SDKs, AWS CLI, or the CreateTransformJob API.

G6e support for Batch Transform is available in US East (N. Virginia), US East (Ohio), US West (Oregon), Asia Pacific (Mumbai), and Asia Pacific (Hyderabad). To learn more, visit the Amazon SageMaker AI product page and see the Batch Transform documentation. For pricing information on these instances, please visit our pricing page.

Categories: general:products/amazon-sagemaker,marketing:marchitecture/artificial-intelligence,marketing:marchitecture/compute,general:products/amazon-ec2,general:products/amazon-sagemaker-deploy

Source: Amazon Web Services

Share This Update