





Mid-level cloud/DevOps role with metro location and known brand but niche HPC/GPU specialization increases selectivity.
HPC, GPU optimization, cluster schedulers and hybrid cloud expertise are highly domain-specific and less transferable.
Multiple explicit mandatory skills and an explicit 5+ years requirement imply strict technical filtering.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Build and operate HPC environments on AWS, Azure, and GCP, including hybrid and multi-cloud architectures for HPC workloads.
Develop Infrastructure as Code using Terraform, CloudFormation, Ansible, and Python; automate provisioning, configuration, CI/CD pipelines, and monitoring of HPC clusters.
Design and manage scalable AI/ML data pipelines, optimize GPU utilization, and implement observability and security solutions for cloud and HPC infrastructure.
5+ years of experience in DevOps and Cloud infrastructure management.
Bachelor's or Master's degree in Computer Science, Engineering, Information Systems, or related field.
Strong Linux system administration skills (RHEL, Rocky Linux, Ubuntu).
Hands-on experience with Python, Bash or Go; Terraform, Ansible, Git, Jenkins/GitHub Actions; container technologies like Docker, Kubernetes, Singularity/Apptainer; and cloud platforms AWS, Azure, or GCP.
Experienced in managing large-scale HPC clusters and distributed computing across on-premises and multi-cloud environments.
Proficient in automating HPC infrastructure deployment and operations with Infrastructure as Code and CI/CD practices.
Comfortable supporting AI/ML workloads with knowledge of AI frameworks (TensorFlow, PyTorch), GPU acceleration, and HPC workload managers (Slurm, LSF, PBS Pro, Kubernetes).