Senior Network Reliability Engineer, Incident Management
SkyloMatch Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Own the 24x7 incident management bridge calls for Sev 1-4 incidents across RAN, 5G Core, Cloud Infrastructure, and OSS domains to ensure rapid incident triage and full-service restoration.
Manage end-to-end incident lifecycle including ticket creation, technical documentation, escalation, and post-incident reviews to meet network availability KPIs and SLA commitments.
Coordinate with internal teams and external vendors/MNO partners during incidents, driving structured communication and continuous procedural improvements.
Minimum Requirements
5-10+ years experience in telecom/wireless operations or network reliability engineering in a production 24x7 environment.
Strong knowledge of telecom networks including RAN (CUSM, eCPRI, PTP/SyncE), 5G Core (AMF, SMF, UPF), Cloud, and OSS to triage and escalate incidents with technical context.
Proficiency with incident management tools (Jira, ServiceNow), observability platforms (Grafana, Prometheus, Loki), Kubernetes operational commands (kubectl), and on-call scheduling tools (PagerDuty or equivalent).
Ability to independently command incident bridge calls for Sev 1-4 incidents, deliver clear incident timelines and manage multiple simultaneous events under pressure.
Ideal Candidate Profile
Experienced in high-pressure, multi-domain telecom incident management with demonstrated leadership in commanding incident response from alert to resolution.
Technically adept with solid operational knowledge across satellite/NTN networks, telecom infrastructure, and cloud-native environments to facilitate precise escalation and problem resolution.
Skilled at cross-functional collaboration with RAN, Core, Cloud, Engineering teams and external partners to maintain network availability and drive continuous improvement initiatives.
