





Senior, niche HA/DR platform role at a known SaaS firm creates moderate applicant density.
Highly specialized HA, DR, networking and cloud skills limit cross-industry transferability.
Explicit 8+ years and mandatory IaC, Kubernetes, networking, HA/DR skills enforce strict filters.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Own the architectural strategy and technical execution for platform high availability, resilience, and disaster recovery lifecycle across cloud and on-premises environments.
Design and implement active-active clustering, global load balancing, and real-time database replication to eliminate single points of failure and optimize uptime.
Lead chaos engineering practices, define resiliency standards, and coordinate multi-region disaster preparedness with cross-functional teams.
8+ years of experience in high availability engineering across cloud (AWS, GCP, or Azure) and on-premises.
Advanced proficiency with Infrastructure as Code tools such as Terraform or Ansible.
Deep technical knowledge of Linux internals, Kubernetes container orchestration, and hybrid network engineering including BGP routing, DNS management, Anycast, and CDNs.
Work Experience Required: Minimum 8 years in relevant HA engineering roles.
Experienced in translating high-level uptime and resiliency requirements into actionable technical roadmaps and execution.
Proven ability to influence and collaborate with cross-functional engineering teams to embed self-healing and automated recovery mechanisms.
Strong strategic execution capabilities combined with hands-on expertise in failover protocols, replication, and resiliency engineering frameworks.