- Adds support for G7 instances (ml.g7.xlarge to ml.g7.48xlarge) in SageMaker AI inference, delivering up to 4.6x faster performance than G6.
- Provides 32 GB GPU memory, 5th‑gen Tensor Cores, up to 700 Gbps EFA networking, and up to 7.6 TB local NVMe storage for large generative‑AI models.
- Instances are available in US East (N. Virginia, Ohio) and US West (Oregon) and can be launched via console, API, or SDK.