Best practices, reference architectures, and examples for distributed AI training and inference on AWS.
-
Updated
Sep 1, 2026 - Shell
Best practices, reference architectures, and examples for distributed AI training and inference on AWS.
Contains example recipes that demonstrate how to build HPC systems using AWS services and solutions.
Deploy your HPC Cluster on AWS in 20min. with just 1-Click.
A curated list of awesome Amazon Web Services (AWS) libraries, open source repos, guides, blogs, and other resources for Academic Researchers new to AWS
Monitoring Dashboard for AWS ParallelCluster AWS ParallelCluster & AWS PCS + Amazon RES virtual desktops
This collection of helper scripts/ and guides for AWS SageMaker HyperPod and ParallelCluster makes it easy to get started with large-scale distributed training on Slurm-based HPC clusters and Kubernetes-based EKS clusters, as well as AI/ML model inference deployment.
This repository contains a collection of AWS ParallelCluster post-install scripts.
Snakemake profile for running on AWS ParallelCluster with the Slurm scheduler
Helping MD customers run workloads on Graviton 3E instances (especially hpc7g) with optimized compiler settings
Seamlessly deploy cryoSPARC in the cloud using AWS CloudFormation and AWS ParallelCluster.
Example of how to run OpenFOAM on a Singularity container through AWS ParallelCluster.
Quickly deploy a multi-user High Performance Computing cluster on AWS using ParallelCluster.
To associate your repository with the parallelcluster topic, visit your repo's landing page and select "manage topics."