Site Reliability Technical Operations Manager
Johnson ControlsMatch Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Manage day-to-day technical operations, L2/L3 support, incident response, and escalations for cloud products across global environments primarily on Azure.
Own service reliability governance including monitoring availability, incident trends, MTTR, and operational risk mitigation.
Drive continuous improvement through incident management, automation, support process standardization, and collaboration with Engineering, Security, and Observability teams.
Minimum Requirements
Minimum 10+ years experience in technical operations, production support, SRE, or cloud reliability engineering.
Strong experience with Azure cloud and familiarity with cloud-native technologies including microservices, Kubernetes, and containers.
Proven expertise in incident, problem, and change management using ITIL-based processes and tools such as Jira and ServiceNow.
Leadership experience managing cloud product support operations in global environments with stakeholder and vendor management.
Ideal Candidate Profile
Experienced in leading operational excellence and reliability practices including SLIs, SLOs, SLAs, and error budget implementation.
Skilled in observability tools such as Grafana, Datadog, ELK, or Azure Monitor with a troubleshooting mindset across cloud platforms, networking, and databases.
Comfortable coordinating across Engineering, SRE, Security, and external partners to improve production stability and lead high-severity incident responses.
