Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Own day-to-day operational reliability and troubleshooting for enterprise network services, including incident and problem management with end-to-end RCA ownership.
Design, implement, and automate production-grade network reliability improvements using Python, Shell, Ansible, and software-defined networking technologies (SD-WAN, SDA).
Drive observability enhancements with dashboards, alerting, and service health metrics, while leveraging enterprise AI tools for operational risk identification and incident triage.
Minimum Requirements
3+ years applied experience with formal training or certification in Site Reliability Engineering concepts.
Strong hands-on skills in enterprise routing/switching and L4–L7 network security components (firewalls, load balancers, proxies).
Proficiency in automation scripting with Python, Shell, and Ansible for production operations.
Experience supporting software-defined networking (e.g., SD-WAN, SDA).
Ideal Candidate Profile
Experienced SRE with demonstrated incident response and problem management leadership including high-quality RCA delivery.
Skilled in integrating AI-assisted workflows within SRE operations with strong validation and data sensitivity awareness.
Comfortable working independently in complex, regulated environments, preferably financial services, with technical depth in Cisco ACI/fabrics and relevant certifications (CCNP preferred).
