AIToday
Amazon AI BlogPublished: May 8, 2026, 01:00 JST1 min read

AWS launches EC2 Capacity Blocks for ML and SageMaker training plans to reserve GPU capacity for short-term workloads

AWS launches EC2 Capacity Blocks for ML and SageMaker training plans to reserve GPU capacity for short-term workloads

3 Key Points

  1. Amazon EC2 Capacity Blocks for ML lets customers reserve GPU capacity for a specific time window, with durations from 1–14 days (in 1-day increments) or 15–182 days (in 7-day increments), at 40-50% lower hourly rates compared to on-demand pricing. For example in US East (N. Virginia), p5.48xlarge costs $34.608/hour with Capacity Blocks versus $55.04/hour on-demand.

  2. SageMaker training plans provide access to reserved GPU capacity within the Amazon SageMaker AI managed environment for training jobs, HyperPod clusters, and inference workloads, priced 70-75% below on-demand rates, with upfront payment required.

  3. Capacity Blocks support selected instance families such as P5, Trn1 and Trn2, but do not cover every GPU instance type and cannot be used with Amazon SageMaker-managed instance types such as ml.p4dn or ml.p5, while SageMaker training plans cover a range of accelerated computing options including the latest NVIDIA GPUs and AWS Trainium accelerators (except G-type instances except G6).

Ask the AI about this article →

Amazon AI BlogRead Original Article

Get AI news like this every morning

For example, today's edition would include:

  • DataAgent launches with $10M to auto-fix Kubernetes faultsSiliconANGLE AI · 48m ago
  • Taoyuan pitches northern AI data center hubDIGITIMES Asia · 48m ago
  • SK Hynix custom HBM boosts inference up to 5.15xDIGITIMES Asia · 48m ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Next articleAMD reports net income up 95% to $1.38 billion in Q1, driven by AI infrastructure demand; projects Q2 revenue of about $11.2 billion, plus or minus $300 million.