Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessMetro location, mid-level SRE role, broad cloud/Kubernetes skillset, and hybrid model create high applicant competition.
Core SRE and cloud skills transfer across industries moderately well but require domain reliability experience, giving medium sensitivity.
Explicit 6+ years requirement, mandatory cloud/Kubernetes/IaC skills and 24/7 on-call make filtering strict and technical.
Job Description
Structured overview of role & requirementsAbout This Role
Own and manage critical and high production incidents end-to-end in a 24/7 environment including leading incident calls and customer communications.
Drive improvements in MTTR, MTTA, alert quality, operational stability, and lead automation initiatives using Go, Python, Shell, or Perl to reduce manual intervention.
Provide technical leadership and mentorship to ProdOps engineers and influence architecture decisions for reliability and operability.
Minimum Requirements
5-8 years experience in Production Operations, SRE, or Cloud Reliability roles.
Proven experience leading major production incidents in customer-facing systems.
Strong background in distributed systems, Kubernetes, and cloud environments (AWS/GCP/Azure).
Must work 24/7 rotational shifts including nights and weekend on-call.
Ideal Candidate Profile
Experienced in large-scale production system support with a focus on incident management and site reliability engineering.
Capable of driving operational maturity and mentoring junior engineers in a hybrid Bangalore location.
Comfortable working in critical incident leadership roles with customer-facing communication responsibilities.
