Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessStrong Tier-1 brand, metro role, broad SRE skillset, and widely sought senior engineering title.
Core SRE and cloud skills transfer well across industries; healthcare preference is only preferred.
Many mandatory cloud, IaC, SRE and incident-response skills required despite no numeric years.
Job Description
Structured overview of role & requirementsAbout This Role
Provide operational support and lead incident management (P1/P2) for critical business applications, ensuring timely resolution and communication.
Develop, maintain, and improve monitoring, alerting, observability, and automation solutions to enhance operational efficiency and service reliability.
Design and deploy AI-powered solutions to improve incident detection, operational workflows, and decision-making at enterprise scale, while mentoring junior engineers.
Minimum Requirements
Hands-on experience with Terraform, GitHub Actions, CI/CD pipelines, and Infrastructure as Code.
Experience with cloud platforms such as Azure and AWS.
Proven experience handling critical P1/P2 incidents in large-scale enterprise environments and solid experience in SRE, Production Support, or Operations Engineering.
Work Experience Required: Not explicitly mentioned in the JD.
Ideal Candidate Profile
Experienced engineer skilled in critical incident management within large-scale, enterprise, cloud-based environments.
Strong background in SRE practices coupled with automation and AI/ML operational solution deployment to improve reliability and efficiency.
Comfortable working with cross-functional teams to drive operational excellence and mentor junior staff in a regulated or complex organizational setting.
