





Tier-1 brand, metro location, mid-level seniority increase applicants, but specialized HPC skills reduce competition.
Specialized HPC, kernel, scheduler, and networking expertise limits transferability across industries.
Explicit 5+ years and mandatory HPC, kernel, scheduler, and storage/network expertise create high filtering.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Own architecture reviews and performance evaluation for new HPC data center clusters across multiple global sites, focusing on compute, storage, networking, and cooling.
Lead benchmarking, profiling, and tuning activities across hardware, OS/kernel, and scheduler layers to optimize HPC infrastructure performance.
Collaborate with platform teams and vendors to assess new hardware and improve observability, capacity planning, and operational readiness for key tapeout milestones.
B.E./B.Tech or M.Tech/M.S. degree.
5+ years of hands-on experience in HPC infrastructure, data center architecture, systems engineering, or senior SRE/platform engineering roles at scale.
Experienced in Linux HPC cluster administration with workload managers (LSF and/or Slurm) and skilled in performance benchmarking and cluster health profiling.
Strong knowledge of data center architecture including compute (CPU/GPU), storage (parallel file systems, NVMe), high-speed networking (InfiniBand, Ethernet, NVLink), and OS/kernel level tuning (NUMA binding, socket affinity, huge pages).
Demonstrated ability to independently drive architecture alignment and provide clear, data-backed tradeoff recommendations to leadership across complex infrastructure programs.
Experienced in multi-site HPC cluster design and operational readiness in fast-paced technology milestones such as tapeouts.
Proficient in scripting (Python, Bash, or Perl) to automate analysis and performance tuning, indicating a hands-on, detail-oriented approach to HPC platform optimization.