





Metro location and broad SRE/cloud skillset increase competition, but seniority and specialization moderate applicant density.
High because role requires deep SRE, cloud, and incident-management expertise, limiting transferability across industries.
High due to explicit 12+ years, 5+ leadership requirement, and mandated SRE/cloud tooling experience.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Lead and manage global Service Reliability Engineering (SRE), Operations, and Production Support teams to ensure availability, performance, and scalability of business-critical platforms.
Own Major Incident Management, post-incident reviews, root cause analysis, and drive corrective/preventive actions to reduce recurring issues and manage crisis communications.
Drive automation, AIOps initiatives, cloud infrastructure reliability (AWS, Azure, GCP), SLA compliance, operational efficiency, and executive-level service delivery governance.
12+ years of IT Operations / Infrastructure experience.
5+ years of leadership in SRE, DevOps, NOC, or Production Support teams.
Experience with cloud platforms (AWS, Azure, GCP), Kubernetes, Linux/Unix, CI/CD, infrastructure as code, scripting (Python/PowerShell/Bash), and ITIL framework.
Bachelor’s degree in Computer Science, Engineering, or related field preferred.
Experienced in managing large-scale, global SRE and operations teams supporting enterprise-level customers.
Proficient in incident and problem management with ability to lead crisis management and executive communications.
Strong background in driving automation, cloud platform modernization, service reliability engineering, and operational governance.