





Tier-1 brand, mid-level generalist SRE role, and popular monitoring skillset increase applicant competition.
SRE skills are broadly transferable across industries, though monitoring tool experience slightly biases fit.
Explicit 2–5 years plus mandatory monitoring, CI/CD and cloud tool experience raises screening rigor.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Accountable for availability, latency, performance, monitoring, and capacity planning of large-scale distributed systems platforms.
Operate, improve, and administer monitoring tools such as OP5/Nagios, Datadog, including alert/event management and synthetic test creation.
Partner with engineering, vendors, and client services to troubleshoot complex issues and deliver technical solutions with limited supervision.
2-5 years of relevant work experience in monitoring toolsets, distributed systems, or site reliability engineering.
Bachelor's degree preferred, but equivalent coursework or extensive professional experience considered.
Expertise with monitoring tools (OP5/Nagios, Datadog) and automation/CI-CD tools (Ansible, Terraform, Puppet, Chef, Jenkins).
Working knowledge of OS management (Windows, Linux), cloud platforms (AWS, Azure, Google), and network technologies (TCP/IP, DNS, SSL, Firewalls).
Experienced in operating and supporting monitoring toolsets in multi time zone, enterprise environments with Agile DevOps practices.
Skilled in troubleshooting complex distributed computing systems and enabling continuous service improvements.
Strong collaborator who can work effectively across DevOps, incident management, problem investigation, and cybersecurity teams.