Home icon

Speed up your cluster procurement time with Amazon SageMaker HyperPod training plans

Machine Learning Blog



AWS introduces SageMaker HyperPod training plans, a new feature designed to help organizations quickly and efficiently procure compute resources for large language model (LLM) training and fine-tuning.

  • Addresses challenges of securing high-performance GPU compute capacity for AI model training
  • Provides predictable access to accelerated compute resources like P4d, P5, and trn2 instances
  • Offers two ways to create training plans: via SageMaker console or AWS CLI
  • Can be used with both SageMaker training jobs and SageMaker HyperPod clusters
  • Allows organizations to search, reserve, and manage compute capacity with simple interfaces

The solution helps organizations overcome compute resource constraints, reduce training cluster procurement wait times, and accelerate AI initiatives by providing flexible, scalable access to high-performance computing resources.



Go to article

The AWS News Feed is currently looking for gold sponsors. If you want to support the AWS community and reach a large audience of AWS professionals, consider sponsoring the AWS News Feed.

Related articles

Dec 4
2024
Amazon SageMaker HyperPod now provides flexible training plans
Dec 4
2024
Meet your training timelines and budgets with new Amazon SageMaker HyperPod flexible training plans
Dec 4
2024
Accelerate foundation model training and fine-tuning with new Amazon SageMaker HyperPod recipes
Aug 11
2025
Amazon SageMaker HyperPod now provides a new cluster setup experience

The AWS News Feed is currently looking for silver sponsors. If you want to support the AWS community and reach a large audience of AWS professionals, consider sponsoring the AWS News Feed.