Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Manage stable, resilient application environments focused on minimizing disruption to Customer & Colleague Journeys (CCJ).
Identify and automate manual tasks, implement observability solutions, and define error budgets to balance risk and reliability.
Lead improvements in release processes, coach colleagues, manage incident responses, and communicate status updates to stakeholders.
Minimum Requirements
Experience Required: Not explicitly mentioned in the JD but role suggests experienced in site reliability or similar.
Strong knowledge in reliability systems thinking and software engineering.
Hands-on experience with AWS tools (EMR, Airflow, S3, EKS, EC2, EMR-Serverless, Lambda, CloudWatch), Spark, Python, shell scripting, DevOps tools like GitLab, GitLab CI, Artifactory, Docker, Prometheus, and Grafana.
Work Location: Chennai, India (Onsite requirement).
Ideal Candidate Profile
Experienced in financial services technology with ability to assess business impact, risk, and connect processes.
Operates using a data-driven, scientific approach to problem solving and reliability engineering.
Capable of engaging and coordinating with diverse stakeholders to manage complex data issues and service management.
