





Known employer and metro location increase competition, but senior 10+ years requirement reduces applicant density.
Deep SRE, cloud, and observability expertise required, limiting cross-industry transferability.
Multiple mandatory SRE skills, 10+ years experience, and cloud/tooling requirements enforce strict shortlisting.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Develop and maintain observability tools, dashboards, and alerts to monitor system health and enhance service delivery.
Collaborate with engineering and product teams to implement best practices in system architecture, capacity planning, automation, and incident response.
Lead initiatives to improve system performance, reduce operational toil through automation, and participate actively in incident management and on-call duties.
Bachelor's degree in Engineering or related technical discipline.
10+ years of experience in Site Reliability Engineering.
Proficiency in observability tools like Splunk/ELK, Datadog, Prometheus, and Grafana.
Experience with scripting languages (Python, Bash, PowerShell) and cloud technologies (GCP, AWS, or Azure).
Experienced in engineering or cloud environments with strong software engineering skills for automation and reliability.
Skilled in distributed system design, CI/CD pipeline management, and configuration management tools such as Terraform and Ansible.
Capable of mentoring junior engineers and influencing technical and business outcomes through strategic collaboration.