AWS Compute Blog
Integrating an Inferencing Pipeline with NVIDIA DeepStream and the G4 Instance Family
Contributed by: Amr Ragab, Business Development Manager, Accelerated Computing, AWS and Kong Zhao, Solution Architect, NVIDIA Corporation AWS continually evolves GPU offerings, striving to showcase how new technical improvements created by AWS partners improve the platform’s performance. One result from AWS’s collaboration with NVIDIA is the recent release of the G4 instance type, a technology […]
Scalable deep learning training using multi-node parallel jobs with AWS Batch and Amazon FSx for Lustre
Contributed by Amr Ragab, HPC Application Consultant, AWS Professional Services How easy is it to take an AWS reference architecture and implement a production solution? At re:Invent 2018, Toyota Research Institute presented their production DL HPC architecture. This was based on a reference architecture for a scalable, deep learning, high performance computing solution, released earlier […]
Amazon ElastiCache performance boost with Amazon EC2 M5 and R5 instances
Contributed by Ruchita Arora, Sr. Product Manager, Allen Farris, Software Dev Engineer, and Itay Maoz, Sr. Software Engineering Manager Earlier this year, Amazon EC2 introduced two exciting new instance families, M5 and R5. These instances are based on the new AWS Nitro system, a combination of dedicated hardware and lightweight hypervisor that aims to deliver […]
Deploying a Burstable and Event-driven HPC Cluster on AWS Using SLURM, Part 2
Contributed by Amr Ragab, HPC Application Consultant, AWS Professional Services In part 1 of this series, you deployed the base components to create the HPC cluster. This unique deployment stands up the SLURM headnode. For every job submitted to the queue, the headnode provisions the needed compute resources to run the job, based on job […]
Deploying a Burstable and Event-driven HPC Cluster on AWS Using SLURM, Part 1
Contributed by Amr Ragab, HPC Application Consultant, AWS Professional Services When you execute high performance computing (HPC) workflows on AWS, you can take advantage of the elasticity and concomitant scale associated with recruiting resources for your computational workloads. AWS offers a variety of services, solutions, and open source tools to deploy, manage, and dynamically destroy compute […]
Improving application performance and reducing costs with Amazon EBS-Optimized Instance burst capability
Contributed by Sooraj Prasannan, Senior Product Manager, Amazon Elastic Block Store In November 2017, Amazon EC2 introduced C5 compute-intensive instances and M5 general-purpose instances. In the first half of 2018, we released EC2 C5d instances and M5d instances by adding high-speed, ultra-low latency local NVMe storage to the EC2 C5 and M5 instance families. EC2 […]
Deploy an 8K HEVC pipeline using Amazon EC2 P3 instances with AWS Batch
Update – April 14, 2020: AWS Elemental MediaConvert now supports 8K UHD video encoding. 8K encoding is available in the MediaConvert on-demand, professional tier, for resolutions up to 8192 x 4320 using HEVC encoding at 10-bit including HDR. To learn more, please visit https://aws.amazon.com/about-aws/whats-new/2019/11/8k-resolution-encoding-now-available-with-aws-elemental-media-convert/. Contributed by Amr Ragab, HPC Application Consultant, AWS Professional Services AWS provides several […]
Building a GPU workstation for visual effects with AWS
Contributed by Mike Owen, Solutions Architect, AWS Thinkbox The elasticity, scalability, and cost effectiveness of the cloud value proposition is attractive to media customers. One of the key design patterns in media and entertainment (M&E) workloads is using the cloud as a content lake and bringing the underlying processes closer without having to synchronize data. […]




