#parallelcluster
some cryoEM validation runs via cryosparc.com software integrated with #AWS parallelcluster HPC and using an FSX/Lustre parallel filesystem built from S3 input cryodata bucket via Lustre/S3 data repository associations. Can't do that on-prem!
June 28, 2025 at 7:22 PM
📰 New article by Brendan Bouffler

What’s the difference between AWS ParallelCluster and AWS Parallel Computing Service?

#AWS #HPC
What’s the difference between AWS ParallelCluster and AWS Parallel Computing Service?
It’s been a year since we announced AWS Parallel Computing Service (PCS). In a way this is the third generation of Slurm-based HPC orchestrators that we’ve brought to you. We’ve learned much from helping customers deploy serious production workloads on AWS ParallelCluster, which itself grew from the foundations layed by CfnCluster – the open-source project [...]
aws.amazon.com
October 28, 2025 at 3:01 PM
I do a lot of #HPC clusters running #AWS #Parallelcluster for computational chemistry, specifically integrating it with the www.schrodinger.com small molecule & molecular dynamics tools.

Wallclock time to create a new on-demand GPU node & scale it in for a job was 4min 9sec. Fastest yet!
February 23, 2025 at 5:31 PM
Used the US long holiday weekend to run #AWS #ParallelCluster build-pipeline and cluster deploy through a logging squid HTTP proxy to document internet destinations used -- this helps set firewall rules for orgs that block egress by default in research VPCs
ParallelCluster, External Destinations & Firewall Rules - BioTeam
When Internet egress is blocked by default inside your cloud environment, what external sites and destinations need to be allowed for AWS Parallelcluster?
bioteam.net
October 14, 2025 at 1:43 PM
I used to post google bait on that other site that would allow people experiencing the same problem I had to find a potential solution.

Lets see if this works on bsky. Gonna post a multi-post thread on the most opaque #AWS #Parallelcluster #HPC debugging hassle I (very occasionally) encounter ...
February 21, 2025 at 7:00 PM
Don't know how many bioinformatics and HPC folks are on this app, but we just released AWS ParallelCluster 3.6.0 which has a slew of good features, including integrated GPU health checks. Check out the release notes here:

https://docs.aws.amazon.com/parallelcluster/latest/ug/document_history.html
May 24, 2023 at 1:44 PM
AWS ParallelCluster is honestly such an incredibly useful tool for large-scale distributed training:
GitHub - aws/aws-parallelcluster: AWS ParallelCluster is an AWS supported Open Source cluster management tool to deploy and manage HPC clusters in the AWS cloud.
AWS ParallelCluster is an AWS supported Open Source cluster management tool to deploy and manage HPC clusters in the AWS cloud. - aws/aws-parallelcluster
buff.ly
December 14, 2024 at 4:25 AM
I miss the "job arrays seem complicated so I just scripted a loop to sbatch 80,000 individual jobs ..." conversations. T

Today most of my requests are due to #aws #parallelcluster auto-scaling failing for quota or "insufficient ec2 capacity" errors -- both error modes not easily visible to users
January 6, 2025 at 1:20 PM
Blog post covering some recent experimentation with AWS ParallelCluster monitoring/alerting via Cloudwatch Subscription Filters wired up to a Slack channel via webhook
AWS Parallelcluster Monitoring - BioTeam
Stop finding out about HPC cluster problems from frustrated users. This post walks through some basic experiments with a Terraform-based approach to proactive AWS ParallelCluster monitoring, using Clo...
bioteam.net
January 26, 2026 at 8:02 PM
SUCCESSFUL HPC!

Note to self: "Cross-account sharing of an #aws #parallelcluster custom AMI hosted in the SharedServices AWS account to all other workload AWS accounts should be fast and simple .."

WRONG. I don't do enough cross-account KMS CMK crypto stuff to nail the IAM/policy bits correctly
November 29, 2024 at 8:38 PM
AWS ParallelCluster 3.16 just fixed that by dropping pcluster-diag right into the AMI.

→ One command for health checks → Clean, structured reports → Ends the guessing game
Less hype, more stability. 🛠️

Link's in the thread👇

Follow @imsampro.bsky.social

#imsampro #AWS #HPC #CloudOps
August 25, 2026 at 7:31 AM
Dive into the full v3.16 release notes right here: aws.amazon.com/about-aws/wh...
AWS ParallelCluster 3.16 adds an on-node diagnostics tool - AWS
Discover more about what's new at AWS with AWS ParallelCluster 3.16 adds an on-node diagnostics tool
aws.amazon.com
August 25, 2026 at 7:31 AM
AWS ParallelCluster 3.15 with support for P6-B300 and Slurm 25.11

AWS ParallelCluster 3.15 adds P6-B300 instances because apparently we haven't milked enough GPU naming schemes yet. "Available at no additional charge" except for, you know, the actual compute that'll cost you a kidney.
March 26, 2026 at 5:03 PM
AWS ParallelCluster 3.14 adds P6e-GB200 and P6-B200 instance types

AWS ParallelCluster 3.14 is now generally available. This release includes P6e-GB200 and P6-B200 instance types, prioritized allocation strategies for optimized instance placement, and N...

#AWS #AwsParallelcluster #AwsGovcloudUs
AWS ParallelCluster 3.14 adds P6e-GB200 and P6-B200 instance types
AWS ParallelCluster 3.14 is now generally available. This release includes P6e-GB200 and P6-B200 instance types, prioritized allocation strategies for optimized instance placement, and NICE DCV support for Amazon Linux 2023. Other features included in this release are support for chef-client log visibility in instance console inside the instance's system log and Amazon Linux 2023 with kernel 6.12. To get started using P6e-GB200 instances with ParallelCluster, follow the tutorial in the ParallelCluster User Guide - https://docs.aws.amazon.com/parallelcluster/latest/ug/support-nvidia-imex-p6e-gb200-instance.html. For more details on the release, review the AWS ParallelCluster 3.14.0 https://github.com/aws/aws-parallelcluster/releases/tag/v3.14.0. ParallelCluster is a fully-supported and maintained open-source cluster management tool that enables R&D customers and their IT administrators to operate high-performance computing (HPC) clusters on AWS. ParallelCluster is designed to automatically and securely provision cloud resources into elastically-scaling HPC clusters capable of running scientific and engineering workloads at scale on AWS. ParallelCluster is available at no additional charge in the AWS Regions listed https://docs.aws.amazon.com/parallelcluster/latest/ug/supported-regions.html, and you pay only for the AWS resources needed to run your applications. To learn more about launching HPC clusters on AWS, visit the ParallelCluster https://docs.aws.amazon.com/parallelcluster/latest/ug/tutorials-v3.html. To start using ParallelCluster, see the installation instructions for ParallelCluster https://docs.aws.amazon.com/parallelcluster/latest/ug/install-pcui-v3.html and https://docs.aws.amazon.com/parallelcluster/latest/ug/install-v3-parallelcluster.html.
aws.amazon.com
September 30, 2025 at 11:05 PM
🆕 AWS ParallelCluster 3.14 adds P6e-GB200 and P6-B200 instances, NICE DCV support, and prioritized allocation strategies. Follow the guide for P6e-GB200. ParallelCluster is a free, open-source HPC cluster management tool for AWS.

#AWS #AwsParallelcluster #AwsGovcloudUs
AWS ParallelCluster 3.14 adds P6e-GB200 and P6-B200 instance types
AWS ParallelCluster 3.14 is now generally available. This release includes P6e-GB200 and P6-B200 instance types, prioritized allocation strategies for optimized instance placement, and NICE DCV support for Amazon Linux 2023. Other features included in this release are support for chef-client log visibility in instance console inside the instance's system log and Amazon Linux 2023 with kernel 6.12. To get started using P6e-GB200 instances with ParallelCluster, follow the tutorial in the ParallelCluster User Guide - Using Amazon EC2 P6e-GB200 UltraServers in AWS ParallelCluster. For more details on the release, review the AWS ParallelCluster 3.14.0 release notes. ParallelCluster is a fully-supported and maintained open-source cluster management tool that enables R&D customers and their IT administrators to operate high-performance computing (HPC) clusters on AWS. ParallelCluster is designed to automatically and securely provision cloud resources into elastically-scaling HPC clusters capable of running scientific and engineering workloads at scale on AWS. ParallelCluster is available at no additional charge in the AWS Regions listed here, and you pay only for the AWS resources needed to run your applications. To learn more about launching HPC clusters on AWS, visit the ParallelCluster User Guide. To start using ParallelCluster, see the installation instructions for ParallelCluster UI and CLI.
aws.amazon.com
September 30, 2025 at 10:40 PM
AWS ParallelCluster 3.16 adds an on-node diagnostics tool

AWS ParallelCluster 3.16 is now generally available with a new on-node diagnostics tool, cluster stability improvements, and an updated HPC and AI/ML software stack.

pcluster-diag is a diagnostics tool built into the ParallelClu...

#AWS
AWS ParallelCluster 3.16 adds an on-node diagnostics tool
AWS ParallelCluster 3.16 is now generally available with a new on-node diagnostics tool, cluster stability improvements, and an updated HPC and AI/ML software stack. pcluster-diag is a diagnostics tool built into the ParallelCluster AMIs that lets you run diagnostic checks on any cluster node with a single command, and get a structured report that makes it easier to identify issues. This release also hardens the cluster lifecycle with more resilient cluster creation, updates, and image builds. The software stack is refreshed, with updated NVIDIA driver, CUDA, EFA installer, and Slurm versions. To get started with pcluster-diag, see https://docs.aws.amazon.com/parallelcluster/latest/ug/troubleshooting-v3-pcluster-diag.html. For more details, review the AWS ParallelCluster 3.16.0 https://github.com/aws/aws-parallelcluster/releases/tag/v3.16.0. AWS ParallelCluster is an open-source cluster management tool that makes it possible for R&D customers and IT administrators to operate high-performance computing (HPC) clusters on AWS. ParallelCluster is designed to automatically and securely provision cloud resources into elastically-scaling HPC clusters capable of running scientific and engineering workloads at scale on AWS. ParallelCluster is available at no additional charge in the AWS Regions listed https://docs.aws.amazon.com/parallelcluster/latest/ug/supported-regions.html, and you pay only for the AWS resources needed to run your applications. To learn more about launching HPC clusters on AWS, visit the ParallelCluster https://docs.aws.amazon.com/parallelcluster/latest/ug/tutorials-v3.html. To start using ParallelCluster, see the installation instructions for ParallelCluster https://docs.aws.amazon.com/parallelcluster/latest/ug/install-pcui-v3.html and https://docs.aws.amazon.com/parallelcluster/latest/ug/install-v3-parallelcluster.html.
aws.amazon.com
August 24, 2026 at 6:05 PM
🆕 AWS ParallelCluster 3.12 now available with custom image build enhancements

#AWS #AwsParallelcluster #AmazonEc2
AWS ParallelCluster 3.12 now available with custom image build enhancements
AWS ParallelCluster 3.12 is now generally available. This release makes it possible to include Lustre and NVIDIA software components in ParallelCluster custom images. Now, you can include ParallelCluster's recommended Nvidia drivers and CUDA libraries in custom images. This update also makes the Lustre client optional to account for scenarios where you may opt for alternative storage solutions. To enable these optional software components when creating custom images, configure the NvidiaSoftware and the LustreClient parameters in the build image configuration file when using the build-image command. For more details on the release, review the AWS ParallelCluster 3.12.0 release notes. AWS ParallelCluster is a fully-supported and maintained open-source cluster management tool that enables R&D customers and their IT administrators to operate high-performance computing (HPC) clusters on AWS. AWS ParallelCluster is designed to automatically and securely provision cloud resources into elastically-scaling HPC clusters capable of running scientific, engineering, and machine-learning (ML/AI) workloads at scale on AWS. AWS ParallelCluster is available at no additional charge in the AWS Regions listed here, and you pay only for the AWS resources needed to run your applications. To learn more about launching HPC clusters on AWS, visit the AWS ParallelCluster User Guide. To start using ParallelCluster, see the installation instructions for ParallelCluster UI and CLI.
aws.amazon.com
December 19, 2024 at 6:23 PM
Interested in how others bootstrap #AWS #parallelcluster #HPC environments.

I tend to hack bash customAction scripts together to create a few SSM parameters that ansible later queries and then we clone an ansible repo and run ansible against localhost inventory to complete final config steps ...
November 29, 2024 at 4:05 PM
Proteros implemented a scalable HPC solution for their Cryo-EM and PX workloads on AWS. Learn how they built secure cryoSPARC on AWS ParallelCluster with hybrid storage yo achieve high performance and low cost. aws.amazon.com/blogs/hpc/ho...
How Proteros accelerates drug discovery by using AWS ParallelCluster | Amazon Web Services
Proteros is a leader in structure-based drug discovery solutions, and supports pharmaceutical, biotechnological, and academic clients with advanced technologies like Cryogenic Electron Microscopy…
aws.amazon.com
November 25, 2025 at 3:21 PM
Proteros implemented a scalable HPC solution for their Cryo-EM and PX workloads on AWS. Learn how they built secure cryoSPARC on AWS ParallelCluster with hybrid storage yo achieve high performance and low cost. aws.amazon.com/blogs/hpc/ho...
How Proteros accelerates drug discovery by using AWS ParallelCluster | Amazon Web Services
Proteros is a leader in structure-based drug discovery solutions, and supports pharmaceutical, biotechnological, and academic clients with advanced technologies like Cryogenic Electron Microscopy…
aws.amazon.com
November 25, 2025 at 3:21 PM
AWS Parallel Computing Service (PCS) is the third generation of Slurm-based HPC orchestrators, building on the foundations of CfnCluster and ParallelCluster. Learn about the key differences between these cloud-enabled HPC solutions. aws.amazon.com/blogs/hpc/th...
October 28, 2025 at 3:02 PM
AWS Parallel Computing Service (PCS) is the third generation of Slurm-based HPC orchestrators, building on the foundations of CfnCluster and ParallelCluster. Learn about the key differences between these cloud-enabled HPC solutions. aws.amazon.com/blogs/hpc/th...
October 28, 2025 at 3:01 PM
📰 New article by Jonathan Davies, Felix Breunig, Robert Meyer, Dominik Schröttinger

How Proteros accelerates drug discovery by using AWS ParallelCluster

#AWS #HPC
How Proteros accelerates drug discovery by using AWS ParallelCluster
Proteros is a leader in structure-based drug discovery solutions, and supports pharmaceutical, biotechnological, and academic clients with advanced technologies like Cryogenic Electron Microscopy (Cryo-EM) and Protein Crystallography (PX). In this blog post, we’ll explore how Proteros implemented an HPC solution that scales with their scientific ambitions. We’ll talk about how they started with a secure [...]
aws.amazon.com
November 25, 2025 at 3:21 PM
Claude code is surprisingly good at sanity checking AWS ParallelCluster HPC config files.

When paired with the MCP that queries instances.vantage.sh it can validate ec2 server config and even provide summarized cost updates on what each compute node config change will cost.
January 22, 2026 at 5:14 PM
If you do #HPC on #AWS via #ParallelCluster then check your versions. A ton of old HPC clusters are gonna fall over in June 2026 when AWS deprecates an old python runtime in Lamda. I've done 3 client upgrades this month and expect more to come ...
March 12, 2026 at 6:26 PM