Home icon

Introducing auto scaling on Amazon SageMaker HyperPod

Machine Learning Blog



AWS has announced auto scaling for Amazon SageMaker HyperPod using Karpenter, an open-source Kubernetes node lifecycle manager. This feature provides managed node automatic scaling for machine learning workloads.

  • Enables just-in-time provisioning of compute resources
  • Supports scaling to zero nodes without maintaining dedicated infrastructure
  • Provides workload-aware node selection and automatic node consolidation
  • Integrates with SageMaker HyperPod's resilience and continuous provisioning capabilities
  • Allows customers to dynamically scale GPU nodes based on real-time demand

The solution can be further enhanced by integrating Kubernetes Event-driven Autoscaling (KEDA) to scale pods based on various metrics, creating a comprehensive auto scaling architecture for machine learning workloads.



Go to article

The AWS News Feed is currently looking for gold sponsors. If you want to support the AWS community and reach a large audience of AWS professionals, consider sponsoring the AWS News Feed.

Related articles

Sep 18
2025
Amazon SageMaker HyperPod now supports autoscaling using Karpenter
Aug 11
2025
Amazon SageMaker HyperPod now provides a new cluster setup experience
Sep 2
2025
Announcing the new cluster creation experience for Amazon SageMaker HyperPod
Aug 22
2025
Amazon SageMaker HyperPod enhances ML infrastructure with scalability and customizability

The AWS News Feed is currently looking for silver sponsors. If you want to support the AWS community and reach a large audience of AWS professionals, consider sponsoring the AWS News Feed.