- G7e instances are now available in Asia Pacific (Seoul), Europe (London), and Asia Pacific (Tokyo) for SageMaker AI inference
- Each instance offers up to 8 NVIDIA RTX PRO 6000 GPUs (96 GB per GPU) and up to 768 GB total GPU memory, delivering up to 2.3× performance over G6e
- The expansion reduces latency for generative AI workloads and supports medium‑to‑large LLM inference (up to 70B parameters) and other high‑memory compute tasks