AWS HPC Blog

Tag: Compute

Resilient HPC and ML on AWS: Running Tightly Coupled Workloads on Spot Instances

This post references AWS ParallelCluster. Check out AWS Parallel Computing Service (AWS PCS), our new managed Slurm service for running HPC and AI workloads on AWS. This post was contributed by Santosh Kumar, Bhagyaraju Kasina, Dr. Sandeep Sovani and Dr. Max Starr Researchers and engineering teams running High Performance Computing (HPC) jobs face a constant […]

Monitoring AWS Parallel Computing Service

This post was contributed by Ronald Hudson and Nate Haynes High Performance Computing (HPC) on AWS demands precise monitoring, like the racing telemetry used by Formula 1 teams to deliver results. Like race engineers tracking car performance, AWS Parallel Computing Service (AWS PCS) administrators must monitor computing metrics in real-time. This vigilance is critical because […]

High Throughput Scheduling for Financial Services with YellowDog HTS on AWS

This post was contributed by Kirill Bogdanov (Pr. Solutions Architect at AWS) and Alan Parry, (CTO at YellowDog). Large-scale compute grids sit at the heart of modern financial services operations. They power overnight batch runs for regulatory risk and prepare traders for the coming day. During trading, the same grids drive intraday ‘value at risk’ […]

Cost-effective and scalable Oxford Nanopore Technologies primary analysis with Nextflow and Amazon EC2 G Instances

This post was contributed by Stefan Dittforth and Michael Mueller Introduction Oxford Nanopore Technologies (ONT) sequencing enhances genome analysis in research and healthcare with its ability to produce long-read sequencing data in real-time. Long reads improve our ability to detect structural variation, resolve repetitive regions, perform haplotype phasing and analyze full-length transcripts, providing a more […]

Scaling life sciences research by deploying AWS ParallelCluster and AWS DataSync

This post references AWS ParallelCluster. Check out AWS Parallel Computing Service (AWS PCS), our new managed Slurm service for running HPC and AI workloads on AWS.   In life sciences research, managing large-scale computational resources and data efficiently is important for success. However, traditional on-premises environments often struggle to meet these requirements effectively. This post […]

How Rivian modernized engineering simulation using AWS

This post references AWS ParallelCluster. Check out AWS Parallel Computing Service (AWS PCS), our new managed Slurm service for running HPC and AI workloads on AWS. This post was contributed by Ameya Kamerkar (Rivian), Vikram Pendyam (Rivian), Abhishek Chauhan (Rivian), Ajay Paknikar (AWS), Sandeep Sovani (AWS) Figure 1. Rivian’s custom Amazon Electric Delivery Vehicle (EDV) […]

AWS re:Invent 2025: Your Complete Guide to High Performance Computing Sessions

This post references AWS ParallelCluster. Check out AWS Parallel Computing Service (AWS PCS), our new managed Slurm service for running HPC and AI workloads on AWS. AWS re:Invent 2025 returns to Las Vegas, Nevada on December 1, uniting AWS builders, customers, partners, and IT professionals from across the globe. This year’s event offers you exclusive […]

Optimizing undersea cables: how Orsted and AWS modeled seabed thermal properties

This post was contributed by Ross Pivovar, Rafał Ołdziejewski, Cindy Xin Qi Lee Offshore wind farms play a critical role in the global transition to renewable energy and clean power generation. But generating electricity is only half the battle—safely and efficiently transporting that power to the grid through undersea cables is equally important. Today, we’ll […]

Announcing expanded support for Custom Slurm Settings in AWS Parallel Computing Service.png

Announcing expanded support for Custom Slurm Settings in AWS Parallel Computing Service

by Brendan Bouffler and Charunethran Panchalam Govindarajan on Permalink Share

Today we’re excited to announce expanded support for custom Slurm settings in AWS Parallel Computing Service (PCS). With this launch, PCS now enables you to configure over 65 Slurm parameters. And for the first time, you can also apply custom settings to queue resources, giving you partition-specific control over scheduling behavior. This release responds directly […]