





Remote role, common SRE title, metro hiring, and broad skillset requirements increase candidate competition.
Core SRE skills are broadly transferable across industries despite domain-specific healthcare context.
Explicit 9+ years and mandatory cloud, Kubernetes, IaC, and monitoring skills make filtering highly strict.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Ensure reliability, scalability, and performance of critical systems and applications by designing and maintaining infrastructure.
Develop automation tools/processes for operations, deployments, and incident response; lead incident management and root cause analysis.
Collaborate with development teams on CI/CD pipeline design and embed reliability in software development lifecycle; mentor junior engineers and participate in on-call rotations.
Bachelor's degree in Computer Science, Engineering, or related field, or equivalent practical experience.
9+ years of experience in Site Reliability Engineering, DevOps, or similar role with large-scale distributed systems.
Proficiency in scripting languages (e.g., Python, Go, Ruby, Bash) and extensive experience with cloud platforms (AWS, Azure, GCP) and containerization (Docker, Kubernetes).
Strong knowledge of infrastructure as code tools (Terraform, Ansible), monitoring/logging tools (Prometheus, Grafana, ELK stack), and CI/CD tools (Jenkins, GitLab CI, Azure DevOps).
Experienced senior-level SRE/DevOps professional with deep expertise in large-scale distributed systems and cloud infrastructure.
Operates with end-to-end ownership of system reliability including incident response, automation, and continual improvement.
Capable mentor and collaborator working closely with development teams to integrate reliability practices into software delivery.