





Metro SRE role with moderate brand and niche AIOps/OTel skills, so medium competition.
High — specialised SRE, observability, OpenTelemetry and AIOps expertise limits cross-industry portability.
High — many mandatory technical skills including OpenTelemetry, AIOps, cloud, IaC, Kubernetes, and programming.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Lead the implementation of AI/ML capabilities like LLMs for automated observability functions including root-cause analysis, anomaly detection, and self-healing infrastructure.
Maintain, optimize, and scale the observability platform ensuring availability, performance, and cost-efficiency while enabling new features with engineering teams.
Develop automation tools using IaC (Terraform) and scripting; contribute to problem management and enforce data quality standards across telemetry pipelines.
Hands-on experience with AI/ML integration in operational workflows, particularly LLMs and predictive analytics.
Proficiency with observability tools (NewRelic preferred, Datadog, Splunk) and OpenTelemetry framework.
Strong coding and scripting skills in Python, Go, or Java; experience with Infrastructure as Code tools like Terraform.
Work Experience Required: Not explicitly mentioned in the JD; Location requirement: Chennai, India; No travel required.
Experienced in embedding AI/AIOps into large-scale observability platforms and leading cross-functional technical strategies in global organizations.
Demonstrates strong expertise in SRE best practices including reliability monitoring frameworks like Google’s Golden Signals, RED, and USE.
Skilled in cloud platforms (AWS/GCP/Azure), containerization (Docker), and orchestration (Kubernetes) with ability to maintain high-availability platforms.