bartek@aws: ~/news
$ whoami
$ AWS Architect · DevOps · Cloud

tag: NVIDIA Blackwell

show all
Thursday, July 23, 2026

AWS Brings G7e GPU Instances to Asia and Europe on SageMaker

Amazon's beefed-up G7e instances are now live in Seoul, London, and Tokyo on SageMaker AI inference. These monsters pack up to 8 NVIDIA RTX PRO 6000 Blackwell GPUs with 96GB each, delivering 2.3x better performance than the previous G6e generation. You can now run massive 70B parameter language models without breaking a sweat or splitting across multiple nodes. Perfect for slashing latency on your generative AI workloads across Asia and Europe.

Wednesday, July 22, 2026

Amazon SageMaker Gets Serious GPU Upgrade with G7 Instances

Amazon SageMaker AI inference now supports G7 instances powered by NVIDIA RTX PRO 4500 Blackwell GPUs, delivering up to 4.6x faster performance than previous G6 models. These beasts pack 32 GB of GPU memory per GPU, 7x better networking, and massive NVMe storage—perfect for running those chunky 7B–30B parameter models without breaking a sweat. No more painful over-provisioning or model quantizing just to fit in memory. Deploy via SageMaker console, API, or SDK and start serving generative AI models like a pro.