Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Lead the reliability, scalability, security, and performance of mission-critical cloud platforms across Azure, GCP, and Kubernetes environments.
Drive AI-powered operations, automation, and self-healing capabilities to enhance efficiency and operational excellence.
Manage major incident response, root cause analysis, problem management, and serve as a technical leader for critical production and customer issues.
Minimum Requirements
Experience with cloud platforms including Azure, GCP, and Kubernetes is mandatory.
Strong skills in observability tools such as Splunk, AppDynamics, and monitoring via logs, metrics, and traces required.
Proficiency in automation, scripting, and Infrastructure as Code (IaC) practices is essential.
Work Experience Required: Not explicitly mentioned in the JD.
Ideal Candidate Profile
Experienced in leading reliability and performance improvements for large-scale, mission-critical cloud platforms.
Comfortable driving AI-driven operational improvements and automation in cloud-native environments.
Proven ability to manage cross-functional collaboration and act as escalation point in high-impact production incidents.
