





Tier-1 brand and Bangalore metro increase competition, while senior niche observability skills moderate density.
Observability and SRE skills transfer across industries but require specialized platform experience.
Explicit 18+ years and specific observability/AIOps/SRE requirements enforce rigid filters.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Define and implement enterprise-wide observability and AIOps strategy to enable proactive detection, intelligent event correlation, and automated incident response.
Lead standardization and instrumentation of telemetry (logs, metrics, traces) aligning with Critical User Journeys (CUJs), SLIs, and SLOs for improved system reliability and operational efficiency.
Collaborate across SRE, application, infrastructure, and service management teams to drive continuous improvement in incident detection, resolution, automation, and self-healing workflows.
18+ years experience in IT Operations, SRE, Observability, or Platform Engineering.
Strong expertise in observability (logs, metrics, traces, distributed tracing), SRE practices (SLIs, SLOs, error budgets), incident management, and automation.
Experience with AIOps including event correlation, alert deduplication, and noise reduction, as well as hands-on with observability platforms and ITSM integrations.
Work Location or Visa Restrictions: Cannot sponsor employment visas or consider candidates on time-limited visa status.
Experienced leader with deep knowledge of distributed systems, cloud environments, telemetry pipelines, and data integration relevant to observability and SRE.
Proven ability to design and implement automation workflows, runbooks, and self-healing mechanisms in complex operational environments.
Skilled at stakeholder management across application, infrastructure, and operations teams to drive observability adoption and continuous operational improvements.