Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Improve and embed non-functional and operational characteristics such as availability, performance, monitoring, security, incident response, and capacity planning across products and services.
Manage day-to-day health of production and non-production environments, responding to incidents and communicating status updates to stakeholders.
Define error budgets, optimize release processes, scale systems through automation, and provide technical guidance and leadership to teams.
Minimum Requirements
Experience working with cloud-native microservices, containerization, Kubernetes workload management, and API management.
Proficiency with Azure, Infrastructure as Code (PowerShell, JSON, Azure Bicep, ARM, Azure DevOps).
Experience with observability tools like Grafana Stack, Log Analytics, AppInsights.
Work Experience Required: Senior-level Site Reliability Engineering experience; specific years not explicitly mentioned.
Ideal Candidate Profile
Experienced senior-level SRE skilled in balancing risk tolerance with reliability through error budgets and incident management.
Strong collaborator capable of engaging with diverse stakeholders including engineers and customers.
Technically adept with DevOps principles, IT Service Management, and automation including orchestration tools and ServiceNow.
