AWS re:Invent 2025 - Train high-performing AI models at scale on AWS (AIM365)
Training large AI models requires significant compute resources and can be time and cost intensive. In this session, learn to optimize and accelerate your model training workloads using AWS's purpose-built infrastructure and tools. We'll dive deep into leveraging services like Amazon SageMaker HyperPod for distributed training at scale and SageMaker fully managed training jobs for cost-effective ML acceleration.