





Senior, niche SRE leadership role attracts moderate competition from experienced candidates in a metro location.
Requires deep SRE and cloud-native expertise, moderately transferable across tech organisations.
Explicit 10–15 year requirement plus extensive mandatory SRE skills makes shortlisting highly strict.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Own the SRE guild charter including defining and enforcing reliability standards, tooling, and best practices across Dev Factory pods.
Lead incident management practices and act as senior escalation point for critical production incidents.
Drive observability strategy, automation, and cross-functional collaboration on infrastructure, capacity, and reliability practices.
10-15 years of experience in related roles.
Bachelor’s or Master’s degree in Computer Science, Engineering, or a related field.
Deep hands-on expertise with observability tools (Prometheus, Grafana, ELK/OpenSearch, Datadog) and incident management frameworks.
Proficiency with Kubernetes, cloud-native infrastructure (AWS, Azure, or GCP), Infrastructure as Code (Terraform, Ansible, Helm), automation scripting (Python, Go, Shell), and CI/CD platforms (Jenkins, GitLab CI, ArgoCD).
Experienced technical leader able to influence multiple teams without formal authority and set cross-pod reliability standards.
Strong background in operationalizing SLIs, SLOs, error budgets, and driving decisions based on them.
Skilled mentor with a focus on reducing operational toil and automating repetitive tasks across a distributed engineering organization.