





Metro location, broad infrastructure skill requirements, and a well-known employer increase competition.
Moderate—NOC and infrastructure skills transfer across industries but require domain-specific operational experience.
Moderate due to explicit 2+ years and multiple mandatory infrastructure, networking, and scripting skills.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Ensure health, stability, and uptime of production systems through real-time monitoring and incident management in a 24/7 operational environment.
Lead incident response efforts including troubleshooting, root cause analysis, and driving resolution during outages and performance issues.
Develop and maintain SOPs, runbooks, and collaborate with engineering teams to improve monitoring, automation, and process consistency.
Minimum 2+ years hands-on experience in Linux/Unix systems administration and network troubleshooting.
Strong knowledge of internet/network protocols (DNS, DHCP, TCP/IP, NTP, SMTP, VPNs, HTTPS, TLS, IPSec) and experience with monitoring tools (Nagios, Datadog, ELK, Splunk, Sumo Logic).
Proficient scripting skills using Shell, Python, or Ruby; familiarity with incident management platforms like PagerDuty, JIRA, or ServiceNow.
Experience with public cloud platforms (preferably AWS), Docker, Kubernetes, CI/CD tools like Jenkins, and Infrastructure-as-Code basics (Terraform).
Experienced in hands-on operational roles managing infrastructure at scale with accountability for uptime and incident resolution.
Comfortable working in a fast-paced, 24/7 support environment, coordinating cross-functionally with DevOps, SRE, Security, and Engineering teams.
Skilled in automation, documentation, and process improvement with a focus on monitoring, incident response, and compliance.