





Medium—senior SRE role with specialized toolset and metro location yields moderate applicant density.
Medium—core SRE skills transfer across industries, though financial-services exposure is preferred.
High—role mandates specific SRE tooling, cloud, Kubernetes, IaC, observability, and leadership skills.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Lead and define Site Reliability Engineering (SRE) best practices across multiple teams ensuring platform reliability, resiliency, and scalability.
Establish and govern Service Level Indicators (SLIs), Service Level Objectives (SLOs), error budgets, and observability solutions with measurable operational metrics.
Drive resiliency testing including chaos engineering, lead incident management with operational excellence, and champion automation and cloud infrastructure reliability.
Proven experience in core SRE practices including system reliability, incident management, automation, and observability.
Hands-on expertise in resiliency testing, chaos engineering, and SLI/SLO/Error Budget frameworks at scale.
Experience with cloud platforms (AWS, Azure, Google Cloud), Docker, Kubernetes, and monitoring tools like Prometheus, Grafana, Datadog, Splunk, or ELK Stack.
Proficiency in scripting/automation tools (Python, Bash, Terraform, Ansible) and managing CI/CD pipelines (Jenkins, GitLab CI/CD, Azure DevOps). Work Experience Required: Not explicitly mentioned in the JD.
Experienced technical leader capable of driving cross-team SRE standards with accountability for service reliability and customer outcomes.
Professional who operates well in complex distributed systems and cloud-native microservices environments, especially with exposure to financial services domain preferred but not mandatory.
Strategic thinker who can influence technical roadmaps and collaborate effectively with Engineering, Security, Product, and Architecture stakeholders.