





Tier-1 brand, metro Bangalore, broad SRE/cloud skillset increase applicant density.
Technical SRE skills transfer across industries, but BFSI domain preference raises sensitivity.
Explicit 10–15 years requirement, domain certifications, and specialized SRE/cloud skills.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Lead enterprise production incident, problem, and change management following ITIL best practices to ensure timely resolution and operational continuity.
Oversee 24x7 production support for L2/L3 application and infrastructure in Linux-based and AWS cloud environments, including disaster recovery and high availability strategies.
Drive service reliability improvements via Root Cause Analysis, automation (CI/CD, Infrastructure as Code), and monitoring using tools like Grafana, ELK, Splunk, CloudWatch, and mentor SRE and support teams to meet SLA/SLO/KPI targets.
10-15+ years IT experience with at least 8+ years in enterprise production support, incident management, cloud infrastructure, and site reliability engineering.
Bachelor's or Master's degree in Computer Science, IT, Engineering, or related field.
Strong experience in Banking & Financial Services domain, specifically Cards & Payments, Mobile Applications, and Cloud-Native Solutions.
Mandatory skills include ITIL-based incident, problem, and change management; AWS cloud services (EC2, S3, Lambda, etc.); Linux environments; Docker, Kubernetes, Jenkins; automation with Python or Bash; monitoring tools like Grafana, ELK, Splunk; AWS certifications preferred.
Deep expertise in production incident and problem management for large-scale AWS cloud environments supporting critical banking & financial services applications.
Experienced leader capable of managing 24x7 production support teams and driving cross-functional coordination during critical incidents.
Hands-on technologist with strong skills in cloud architecture, automation, monitoring, and continuous improvement initiatives focused on reducing MTTR and improving reliability.