Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Lead and own reliability outcomes (availability, performance, recoverability) for mission-critical network services at JPMorgan Chase.
Drive major incident response including coordination, mitigation, coaching, and post-incident learning for medium to large-sized products.
Architect and deliver automation frameworks (Python, Shell, Ansible) and provide deep technical leadership in SD-WAN, routing, switching, firewalls, load balancers, and proxies.
Minimum Requirements
5+ years applied site reliability engineering experience with formal training or certification in SRE concepts.
Extensive experience operating and engineering large-scale networks with strong troubleshooting skills.
Proven leadership in major incident management including coordination and RCA with systemic remediation.
Advanced automation skills in Python, Shell, and Ansible with demonstrable toil reduction and reliability gains.
Ideal Candidate Profile
Experienced in technical leadership roles involving complex network infrastructure including SD-WAN, Cisco ACI/Fabrics, and optionally VMware NSX.
Strong skill set in embedding SRE practices to improve observability, operational readiness, and resilience in a financial institution environment.
Ability to utilize enterprise-authorized AI tools responsibly to enhance incident investigation and operational workflows with strict validation and security control mindset.
