





Tier-1 employer, remote role, mid-level SRE title, and broad toolset requirements increase candidate competition.
Core SRE skills are transferable across industries, though cybersecurity and AI-observability requirements raise domain specificity.
Explicit 5+ years plus mandatory CI/CD, Kubernetes, observability, and programming skills make screening highly strict.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Build and maintain scalable, reliable internal developer platform and CI/CD tools to improve engineering productivity.
Implement automation, observability, and incident response practices including AI integration for proactive system resilience.
Design enterprise-scale highly available services and drive production readiness and reliability improvement initiatives.
5+ years experience in large-scale production environments with SRE or related reliability focus.
Hands-on expertise in CI/CD tools (Bazel, Github Actions, Jenkins), IaC tools (Ansible, Terraform, etc.), and source code management (GitHub, GitLab, Bitbucket).
Experience with Kubernetes deployment and observability tools (Prometheus, Grafana, Datadog, etc.).
Work Experience Required: 5+ years in large-scale production environment.
Experienced Site Reliability Engineer with strong software development and observability expertise at enterprise scale.
Ability to integrate AI-enabled capabilities into observability and operational workflows for automation and self-healing.
Comfortable working in fast-paced, remote and local teams with a security-first mindset and strong programming skills in Python, Go, TypeScript, Java, or C#.