





Popular SRE role, Bangalore metro, broad cloud/Kubernetes skillset and mid-level seniority, causing high competition.
Cloud-native SRE skills transfer across industries but often require platform-specific and enterprise-domain experience.
Explicit 6+ years requirement plus mandatory SRE/cloud/Kubernetes and programming skills increases shortlisting strictness.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Lead onboarding and adoption of reliability standards across services and teams, ensuring adherence to Service Level Objectives (SLOs) and Service Level Agreements (SLAs).
Design, implement, and maintain scalable, resilient, and secure cloud-native infrastructure and monitoring, focusing on incident detection, response, and automation with Python and AI tools.
Drive incident management including root cause analysis and participate in on-call rotations providing 24/7 support for production systems.
6+ years of experience in Site Reliability Engineering or managing infrastructure and services at scale.
Bachelor’s or Master’s degree in Computer Science or equivalent.
Experience managing Hadoop and Kubernetes infrastructure (or equivalent experience) and advanced knowledge of Linux, Networking, and Containers.
Proficiency in at least one high-level programming language (Python, GoLang etc.) and fluency in English.
Strong experience implementing and improving SLOs, SLIs, SLAs aligned with business needs in complex distributed systems.
Proven ability to lead incident response processes and automation of operational tasks, leveraging programming and AI-assisted development tools.
Experience working cross-functionally with development teams to enhance system reliability, scalability, and performance in cloud environments especially Kubernetes and AWS.