Amazon SageMaker AI now supports G7e instances in Asia Pacific (Seoul and Tokyo) and Europe (London). These instances utilize up to eight NVIDIA RTX PRO 6000 Blackwell GPUs and 5th Gen Intel Xeon processors to deliver 2.3x better inference performance than G6e. The expansion allows for lower latency deployment of generative AI workloads closer to users in these regions, supporting models up to 70B parameters.
- G7e instances are now available in Seoul, London, and Tokyo on SageMaker.
- Performance improved up to 2.3x over G6e using Blackwell GPUs and Xeon processors.
- Each instance offers up to 768 GB total GPU memory for large model serving.
- Supports inference endpoints for generative AI models up to 70B parameters.
- Elastic Fabric Adapter provides up to 1,600 Gbps networking bandwidth.