





Metro senior SRE with broad cloud/Kubernetes skills and known employer increases competition.
Platform SRE skills are transferrable across SaaS and cloud companies but expect moderate domain specificity.
Explicit 8–12 years and mandatory AWS, EKS, IaC, and language skills make filters strict.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Design, build, and operate highly reliable, scalable multi-tenant SaaS cloud infrastructure with automation and operational workflows.
Develop internal tools and observability systems, define and track SLOs, and lead incident management and resilience efforts for production systems.
Collaborate with development teams to improve system design, security, scalability, and ensure global 24x7 production support.
8–12 years of experience in Site Reliability Engineering, DevOps, Cloud Infrastructure, or related platform engineering role.
Strong programming skills in Python or Go; deep AWS and Kubernetes (EKS) expertise including Docker, Helm, and networking.
Experience with Infrastructure as Code tools (Terraform, Terragrunt), observability platforms (Datadog, Prometheus, CloudWatch) and incident management.
Work Experience Required: 8–12 years explicit; Degree: Bachelor's or equivalent in Computer Science or related field; Certifications: AWS or Kubernetes preferred but not mandatory.
Experienced engineer with proven ability to design, build, and operate large-scale, fault-tolerant SaaS platform infrastructure under DevOps/SRE practices.
Expert in automation and system resilience, comfortable leading incident response and driving platform feature enhancements.
Strong collaborator who can mentor others, influence cross-functional teams, and leverage AI and data-driven methods to enhance platform operations.