Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Own troubleshooting and reliability improvements for network services including incident and problem management with end-to-end RCA ownership.
Design and implement production-grade automation and software-defined networking solutions using Python, Shell, Ansible, SD-WAN, SDA, and enterprise networking components.
Drive observability, alert quality, and reliability enhancements partnering with development and platform teams, while leveraging AI capabilities responsibly to support SRE workflows.
Minimum Requirements
3+ years applied experience in site reliability engineering with formal training or certification on SRE concepts.
Strong hands-on expertise in enterprise routing/switching and security/L4–L7 network components.
Proficient automation skills using Python, Shell scripting, and Ansible in production environments.
Work Experience Required: 3+ years in relevant SRE/network roles.
Ideal Candidate Profile
Experienced with complex network operational ownership, including incident response, major incident management, and problem management in enterprise environments.
Skilled in designing and automating highly reliable network services with an SRE mindset focused on NFRs, risk analysis, and continuous improvement.
Familiar with software-defined networking technologies and integration of AI-assisted operational tools, particularly within regulated or financial institutions.
