





Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Tier-1 brand and metro location increase applicants, but niche AI/HPC specialization moderates competition.
Role requires specialized AI/HPC GPU infrastructure skills, limiting easy cross-industry transferability.
Explicit 8+ years, 3+ years AI/HPC and many mandatory platform skills make filters highly stringent.
Manage and optimize HPE’s next-generation AI infrastructure platforms including PCAI and AI Factory environments for operational stability and lifecycle management.
Administer and monitor compute nodes, GPU clusters, container platforms, and storage solutions ensuring uptime, performance, and compliance.
Drive automation, incident management, continuous improvement, and knowledge enablement supporting large-scale AI and HPC workloads.
8+ years of IT infrastructure administration experience with at least 3 years in AI/HPC or GPU-based environments.
Bachelor’s or Master’s degree in Computer Science, IT, or equivalent field.
Hands-on expertise with HPE PCAI & AI Factory solutions, NVIDIA AI Enterprise, container orchestration (Kubernetes, Rancher Harvester), and automation tools (Ansible, AWX).
Work Location: Hybrid with an average 2 days per week onsite at HPE office.
Experienced in operating enterprise-grade AI/HPC infrastructure focusing on platform stability, lifecycle management, and automation in containerized and bare metal environments.
Strong technical proficiency with HPE Ezmeral, NVIDIA GPU stack, AI/ML platforms (TensorFlow, PyTorch, Kubeflow) and storage solutions (VAST, WEKA, Alletra).
Capable of leading operational improvements, root cause analysis, and mentoring support teams in a global enterprise IT services organization.