





High — Tier-1 brand, metro location, mid-level generalist SRE role with broad, popular skill requirements.
Medium — skills transferable across tech companies but specialized to reliability and operations contexts.
Medium — specific SRE, cloud, Kubernetes and automation skills required but no explicit years mandated.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Lead incident management as Incident Commander, ensuring resolution of major incidents and effective communication with leadership and partner teams.
Continuously monitor health of critical services to proactively identify and resolve potential issues and collaborate with teams to resolve recurring problems and onboard new alerts.
Enhance automation, monitoring tools, and develop solutions with Architecture, Engineering, and Operations teams to improve site availability and reliability.
Experience in large-scale internet/server environments or equivalent highly-scaled enterprise server environments.
Strong technical triage, troubleshooting, and crisis-level incident management skills with demonstrated leadership in incident handling.
Proficiency with automation programming using languages/technologies like GO, Python, Java, NodeJS, Docker, Kubernetes.
Work location requirement: day shift fixed schedule working 10-hour shifts for four days consecutively; no on-call responsibilities; shift in BanTeam (location not explicitly specified).
Experienced in managing and resolving high-severity incidents within complex, multi-tier and cloud computing environments.
Skilled in collaborating cross-functionally with architecture, engineering, and operations groups to develop effective reliability and automation solutions.
Comfortable working fixed shifts focused on operational reliability without on-call duties, indicating aptitude for predictable schedule and consistent incident oversight.