





Mid-level SRE in metro with broad cloud/Kubernetes skills and recognizable startup brand increases candidate competition.
Technical SRE skills are transferable across industries but require cloud/Kubernetes domain experience.
Explicit 3-5 years and multiple mandatory cloud, Kubernetes, automation, and security skills make shortlisting strict.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Own the reliability, performance, and efficiency of critical services within the platform.
Design and implement automation to reduce operational toil and improve system resilience.
Lead incident management, root cause analysis, and mentor junior SRE engineers while optimizing AWS and GCP infrastructure including Kubernetes orchestration.
3-5 years experience in Site Reliability Engineering, DevOps, or similar role with production systems focus.
Proficiency in Python or Go for complex automation tasks.
Strong experience with AWS and/or GCP cloud platforms and Kubernetes (including tools like ArgoCD, Helm/Kustomize).
Knowledge of cloud security principles, infrastructure as code (Terraform or Ansible), and monitoring/observability tools (Prometheus, Grafana, ELK).
Experienced in managing large-scale distributed systems reliability and incident resolution.
Skilled in cloud infrastructure cost optimization and security-focused operational workflows.
Capable of influencing architecture for scalability and reliability and mentoring junior team members in complex, high-scale environments.