





Tier-1 employer and metro location increase applicant density, though role seniority and specialization moderate it.
Requires deep SRE, cloud, observability, and platform experience, limiting cross-industry transferability.
Explicit 10+ years, SRE certification, cloud/container, and tool mandates make hiring filters stringent.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Lead design, development, and implementation of core observability services including metrics pipelines and log aggregation to boost reliability and operational efficiency.
Manage critical incident management improvements and enhance the end-to-end software development lifecycle across multiple teams within the Card Site Reliability Engineering function.
Leverage enterprise-authorized AI capabilities for complex incident analysis and reliability decision-making while ensuring data sensitivity, security, and auditability compliance.
10+ years of applied Site Reliability Engineering experience with formal training or certification in SRE concepts.
Proficiency in at least one programming language (Python, Java/Spring Boot) and experience with cloud-native (AWS) instrumentation and streaming data platforms.
Experience with CI/CD tools (Jenkins, GitLab, Terraform) and container orchestration technologies (ECS, Kubernetes, Docker).
Demonstrated use of enterprise-authorized AI capabilities for reliability engineering workflows with strong validation and data sensitivity awareness.
Experienced leader capable of managing multiple architects and designers while integrating cross-functional team feedback to drive large projects and system design.
Skilled in defining SLOs and error budgets, championing observability as a first-class concern in the software development lifecycle.
Able to influence technology strategy and operational policies with a commercial mindset enabling innovative, scalable solutions in a complex financial services environment.