





Niche senior observability role at lesser-known firm in Chennai reduces applicant competition.
Core observability and SRE skills are transferable but require specific tools, cloud, and AI experience.
Explicit 8+ years, 3+ years people management, and mandatory observability and cloud tool experience.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Lead and mentor a team of observability engineers to manage monitoring, logging, tracing, and dashboarding solutions across enterprise platforms.
Define and execute observability and AI-driven operations strategy, including adoption of AI/Agentic AI for autonomous incident management and operational efficiency.
Establish KPIs, SLAs, drive reliability, incident response readiness, and collaborate cross-functionally with SRE, cloud, security, and engineering teams.
Bachelor's degree in Computer Science, Engineering, or related field.
Minimum 8 years experience in infrastructure, operations, SRE, platform engineering, or observability domains.
At least 3 years experience in people management leading technical teams.
Experience with observability tools (e.g., Splunk, Datadog, AppDynamics, Grafana, Prometheus), cloud environments (Azure & GCP), and AI-powered observability solutions including Microsoft Copilot and Generative AI.
Experienced leader skilled at managing and scaling observability teams supporting enterprise-scale platforms and services.
Proficient in integrating AI, Agentic AI, and Microsoft Copilot technologies to automate incident management and improve operational workflows.
Strong collaborative operating style partnering with SRE, engineering, cloud, and security teams to enhance service reliability and monitoring capabilities.