





Tier-1 brand and metro location but senior specialized observability skillset lowers candidate density.
Requires specialized observability and infra skills, though expertise is transferable across industries.
Explicit 8+ years requirement and mandatory monitoring, Kafka, Linux, and IaC expertise increases screening rigor.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Lead design, implementation, and maintenance of monitoring strategies across multiple platforms including Linux, Storage, Applications, and Networking.
Define and establish KPI/metrics standards, SLAs, SLOs, and alert thresholds to ensure consistent, measurable observability aligned with business needs.
Architect and deploy sophisticated monitoring solutions incorporating Prometheus, Grafana, Nagios, SolarWinds, Kafka-based event streaming, and AI-driven anomaly detection.
Minimum 8+ years professional experience in monitoring, observability, systems engineering, or DevOps roles.
Proven technical leadership and mentoring experience in monitoring or infrastructure.
Strong hands-on skills with Prometheus, Grafana, Nagios, SolarWinds, Kafka, Linux system administration and bash scripting.
Experience with AI/ML-based monitoring or anomaly detection, and knowledge of metrics collection and time-series data management.
Experienced in defining and driving organization-wide monitoring strategy with measurable impact on system reliability and user experience.
Technical leader comfortable balancing high-level strategic thinkers and hands-on implementation, including mentoring and communication with diverse stakeholders.
Strong domain expertise in real-time event processing, monitoring infrastructure optimization, and emerging observability technologies such as AI monitoring and Infrastructure-as-Code.