Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Ensure high availability, performance, and reliability of production enterprise platforms and applications.
Monitor system health using observability tools and define SLIs, SLOs, and SLAs to measure system performance.
Lead/support incident management, root cause analysis, and automate operational tasks including Infrastructure as Code and CI/CD pipeline improvements.
Minimum Requirements
Experience with cloud platforms like Azure.
Hands-on skills with monitoring tools such as Prometheus, Grafana, Splunk, ELK, or Datadog.
Proficiency in containers and orchestration including Docker and Kubernetes.
Work Experience Required: Not explicitly mentioned in the JD.
Ideal Candidate Profile
Experienced in working closely with engineering and DevOps teams in a production environment.
Strong technical ability in cloud infrastructure automation and CI/CD pipeline management.
Familiar with operational incident response and continuous improvement of large-scale distributed systems.
