+
+

Related Products

  • RunPod
    205 Ratings
    Visit Website
  • Vertex AI
    827 Ratings
    Visit Website
  • Google AI Studio
    11 Ratings
    Visit Website
  • Google Compute Engine
    1,155 Ratings
    Visit Website
  • LM-Kit.NET
    24 Ratings
    Visit Website
  • Ango Hub
    15 Ratings
    Visit Website
  • Dragonfly
    16 Ratings
    Visit Website
  • Gr4vy
    5 Ratings
    Visit Website
  • KrakenD
    71 Ratings
    Visit Website
  • StackAI
    49 Ratings
    Visit Website

About

Amazon EC2 Capacity Blocks for ML enable you to reserve accelerated compute instances in Amazon EC2 UltraClusters for your machine learning workloads. This service supports Amazon EC2 P5en, P5e, P5, and P4d instances, powered by NVIDIA H200, H100, and A100 Tensor Core GPUs, respectively, as well as Trn2 and Trn1 instances powered by AWS Trainium. You can reserve these instances for up to six months in cluster sizes ranging from one to 64 instances (512 GPUs or 1,024 Trainium chips), providing flexibility for various ML workloads. Reservations can be made up to eight weeks in advance. By colocating in Amazon EC2 UltraClusters, Capacity Blocks offer low-latency, high-throughput network connectivity, facilitating efficient distributed training. This setup ensures predictable access to high-performance computing resources, allowing you to plan ML development confidently, run experiments, build prototypes, and accommodate future surges in demand for ML applications.

About

Amazon EC2 UltraClusters enable you to scale to thousands of GPUs or purpose-built machine learning accelerators, such as AWS Trainium, providing on-demand access to supercomputing-class performance. They democratize supercomputing for ML, generative AI, and high-performance computing developers through a simple pay-as-you-go model without setup or maintenance costs. UltraClusters consist of thousands of accelerated EC2 instances co-located in a given AWS Availability Zone, interconnected using Elastic Fabric Adapter (EFA) networking in a petabit-scale nonblocking network. This architecture offers high-performance networking and access to Amazon FSx for Lustre, a fully managed shared storage built on a high-performance parallel file system, enabling rapid processing of massive datasets with sub-millisecond latencies. EC2 UltraClusters provide scale-out capabilities for distributed ML training and tightly coupled HPC workloads, reducing training times.

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Audience

Companies in search of a solution to get scalable access to high-performance compute instances for their machine learning training and inference workloads

Audience

Developers and researchers seeking a solution to improve their machine learning, generative AI, and high-performance computing application operations

Support

Phone Support
24/7 Live Support
Online

Support

Phone Support
24/7 Live Support
Online

API

Offers API

API

Offers API

Screenshots and Videos

Screenshots and Videos

Pricing

No information available.
Free Version
Free Trial

Pricing

No information available.
Free Version
Free Trial

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Training

Documentation
Webinars
Live Online
In Person

Training

Documentation
Webinars
Live Online
In Person

Company Information

Amazon
Founded: 1994
United States
aws.amazon.com/ec2/capacityblocks/

Company Information

Amazon
Founded: 1994
United States
aws.amazon.com/ec2/ultraclusters/

Alternatives

Alternatives

AWS Neuron

AWS Neuron

Amazon Web Services

Categories

Categories

Integrations

AWS Neuron
AWS Nitro System
AWS Trainium
Amazon EC2
Amazon EC2 G5 Instances
Amazon EC2 Inf1 Instances
Amazon EC2 P4 Instances
Amazon EC2 P5 Instances
Amazon EC2 Trn1 Instances
Amazon EC2 Trn2 Instances
Amazon EKS
Amazon Elastic Container Service (Amazon ECS)
Amazon SageMaker
Amazon Web Services (AWS)
PyTorch
TensorFlow
Amazon EC2 Auto Scaling
Amazon EC2 Capacity Blocks for ML
Amazon FSx
Greenovative

Integrations

AWS Neuron
AWS Nitro System
AWS Trainium
Amazon EC2
Amazon EC2 G5 Instances
Amazon EC2 Inf1 Instances
Amazon EC2 P4 Instances
Amazon EC2 P5 Instances
Amazon EC2 Trn1 Instances
Amazon EC2 Trn2 Instances
Amazon EKS
Amazon Elastic Container Service (Amazon ECS)
Amazon SageMaker
Amazon Web Services (AWS)
PyTorch
TensorFlow
Amazon EC2 Auto Scaling
Amazon EC2 Capacity Blocks for ML
Amazon FSx
Greenovative
Claim Amazon EC2 Capacity Blocks for ML and update features and information
Claim Amazon EC2 Capacity Blocks for ML and update features and information
Claim Amazon EC2 UltraClusters and update features and information
Claim Amazon EC2 UltraClusters and update features and information