





Metro location, recognizable multinational brand, and popular SRE role increase applicant competition.
Core SRE skills are widely transferable across industries despite insurance context.
Explicit 7+ years plus mandatory SRE, Datadog, cloud and automation expertise.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Own and improve SLI/SLO/SLA definitions and error budget management for critical enterprise services with executive reporting.
Drive operational excellence including incident/problem/change management aligned to ITIL, reducing toil through automation, enhancing resilience, observability (Datadog), and release reliability.
Partner with engineering, architecture, and DevOps teams to design highly available, fault-tolerant systems and support cloud migration with reliability-by-design principles.
7+ years in SRE, Production Operations, Platform Engineering, or DevOps supporting enterprise applications.
Proven experience in SLA/SLO/SLI implementation, incident response, automation scripting/programming (Python/Go/Java/PowerShell), and hands-on Datadog monitoring/APM.
Working knowledge of cloud platforms (AWS/Azure/GCP) and modern DevOps/CI/CD practices.
Role based in Mumbai with hybrid in-office requirement of at least three days per week.
Experienced leader in implementing and scaling SLO/error budget programs across multiple teams and services.
Strong system design and reliability mindset with expertise in automation, self-healing systems, and observability tooling.
Comfortable managing ITIL-aligned operational processes and collaborating cross-functionally to improve platform resilience and release quality.