





Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Senior, niche monitoring role with specific SolarWinds/Datadog requirements reduces applicant competition.
Core infrastructure and monitoring skills are transferable across industries, so moderate background sensitivity.
Explicit 12+ years plus mandatory tooling, ITIL, and team leadership requirements make shortlisting strict.
Own and lead end-to-end Incident, Problem, and Change Management processes to ensure service restoration within SLA and operational risk mitigation.
Manage infrastructure monitoring and observability platforms (SolarWinds, Datadog) across Network, Servers, Cloud, Applications, and Middleware to improve proactive detection and service visibility.
Lead a team of 10–15 engineers driving operational governance, capacity and availability management, vendor escalations, continuous service improvement, automation, and performance analysis.
Minimum 12 years of experience in IT Operations, Infrastructure Services, Monitoring, or Service Management.
Strong experience with Incident, Problem, Change, and Major Incident Management, including ITIL best practices.
Hands-on knowledge of enterprise monitoring tools: SolarWinds, Datadog, and observability concepts (Metrics, Logs, Traces, APM).
Proven experience managing teams of 10–15 members and driving operational transformations.
Experienced leader comfortable managing high-pressure, customer-facing escalations and interfacing with executives, vendors, and stakeholders.
Strategic operator with skills in optimizing monitoring strategies, SLI/SLO/SLA management, alert rationalization, and continuous service improvement.
Demonstrated ability to mentor and develop teams while overseeing enterprise-scale production operations across multi-cloud and hybrid infrastructure environments.