NVIDIA A100 Marks Dawn of Next Decade in Accelerated Cloud Computing
November 3, 2020 | NVIDIA NewsroomEstimated reading time: 1 minute
Amazon Web Services’ first GPU instance debuted 10 years ago, with the NVIDIA M2050. At that time, CUDA-based applications were focused primarily on accelerating scientific simulations, with the rise of AI and deep learning still a ways off.
Since then, AWS has added to its stable of cloud GPU instances, which has included the K80 (p2), K520 (g3), M60 (g4), V100 (p3/p3dn) and T4 (g4).
With its new P4d instance generally available today, AWS is paving the way for another bold decade of accelerated computing powered with the latest NVIDIA A100 Tensor Core GPU.
The P4d instance delivers AWS’s highest performance, most cost-effective GPU-based platform for machine learning training and high performance computing applications. The instances reduce the time to train machine learning models by up to 3x with FP16 and up to 6x with TF32 compared to the default FP32 precision.
They also provide exceptional inference performance. NVIDIA A100 GPUs just last month swept the MLPerf Inference benchmarks — providing up to 237x faster performance than CPUs.
Each P4d instance features eight NVIDIA A100 GPUs and, with AWS UltraClusters, customers can get on-demand and scalable access to over 4,000 GPUs at a time using AWS’s Elastic Fabric Adaptor (EFA) and scalable, high-performant storage with Amazon FSx. P4d offers 400Gbps networking and uses NVIDIA technologies such as NVLink, NVSwitch, NCCL and GPUDirect RDMA to further accelerate deep learning training workloads. NVIDIA GPUDirect RDMA on EFA ensures low-latency networking by passing data from GPU to GPU between servers without having to pass through the CPU and system memory.
In addition, the P4d instance is supported in many AWS services, including Amazon Elastic Container Services, Amazon Elastic Kubernetes Service, AWS ParallelCluster and Amazon SageMaker. P4d can also leverage all the optimized, containerized software available from NGC, including HPC applications, AI frameworks, pre-trained models, Helm charts and inference software like TensorRT and Triton Inference Server.
P4d instances are now available in US East and West, and coming to additional regions soon. The instances can be purchased as On-Demand, with Savings Plans, with Reserved Instances, or as Spot Instances.
The first decade of GPU cloud computing has brought over 100 exaflops of AI compute to the market. With the arrival of the Amazon EC2 P4d instance powered by NVIDIA A100 GPUs, the next decade of GPU cloud computing is off to a great start.
Suggested Items
Intel Announces New Program for AI PC Software Developers and Hardware Vendors
03/27/2024 | Intel CorporationIntel Corporation announced the creation of two new artificial intelligence (AI) initiatives as part of the AI PC Acceleration Program: the AI PC Developer Program and the addition of independent hardware vendors to the program.
SEMI 3D & Systems Summit To Spotlight Trends In Hybrid Bonding, Chiplet Design And Environmental Sustainability
03/26/2024 | SEMILeading experts in 3D integration and systems for semiconductor manufacturing applications will gather at the annual SEMI 3D & Systems Summit, 12-14 June, 2024,
Ventec to Launch New Bondply Dielectrics and Value-Added Services at IPC APEX EXPO 2024
03/26/2024 | Ventec International GroupVentec International Group is to reveal new products for advanced signal integrity and thermal performance, and introduce services, during IPC APEX EXPO 2024, April 9-11 on booth # 4309.
RTX's Raytheon Lower Tier Air, Missile Defense Sensor Detects and Engages Complex Target
03/25/2024 | RTXRaytheon, an RTX business, announced that its Lower Tier Air and Missile Defense Sensor, or LTAMDS, continues to advance through its U.S. Army test program with another successful live-fire event. Military leaders from seven nations were on-site to witness the radar's capabilities and performance first-hand.
IMAPS Wrap-up: AI, Chiplets, and 3D Cube Architecture
03/22/2024 | Marcy LaRont, PCB007 MagazineThe International Microelectronics Assembly and Packaging Society, IMAPS, held its 20th Device Packaging Expo and Conference this past week in Fountain Hills, Arizona, followed immediately by a ‘Workshop on Advanced Packaging for Medical Electronics’ that continued through the remainder of Thursday. Fortunate to find myself in Texas earlier in the week, I made it for the last day of the IMAPS event and attended two excellent keynote presentations by AMD and Intel, respectively. Here are some highlights.