





Metro location, known employer, and in-demand SRE skills yield medium competition.
Core SRE skills transfer across industries, but multi-tenant SaaS and platform specifics create medium domain bias.
Explicit 8–12 years plus mandatory AWS, EKS, Python/Go, IaC, and observability requirements increase strictness to high.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Design, build, and operate reliable, scalable infrastructure and automation for a multi-tenant SaaS platform using EKS, AWS, and cloud technologies.
Develop and maintain observability systems, lead incident management, and drive improvements toward self-healing and reduced operational toil.
Collaborate across teams to enhance platform security, reliability, and production readiness; participate in 24x7 follow-the-sun on-call support.
8–12 years of experience in Site Reliability Engineering, DevOps, Cloud Infrastructure, or related platform engineering roles.
Strong programming skills in Python or Go, with deep experience in AWS and Kubernetes (EKS), including related tools like Docker, Helm, and Infrastructure as Code (Terraform or similar).
Experience with observability tools (Datadog, Prometheus, etc.), incident management, and production support in microservices environments.
Knowledge of security and identity frameworks (SSO, SAML, OAuth) and AWS IAM; Bachelor’s degree in Computer Science or equivalent experience.
Experienced in operating large-scale SaaS platforms with hands-on skills in cloud-native infrastructure and automation workflows.
Able to lead reliability and observability engineering efforts and participate effectively in follow-the-sun on-call rotations.
Collaborative and technically influential, able to mentor others, implement secure and compliant platform services, and drive automation to reduce manual effort.