Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessSenior SRE in metro with broad cloud, Kubernetes, and observability requirements attracts many qualified candidates.
Core SRE skills transfer across industries, but financial services and regulatory expectations increase fit sensitivity.
Explicit 15+ years plus deep SRE, cloud, automation and observability requirements create strict shortlisting filters.
Job Description
Structured overview of role & requirementsAbout This Role
Own and drive reliability strategy including defining and governing SLIs, SLOs, and Error Budgets for business-critical products.
Lead major incident management, post-incident reviews, and operational improvements to increase MTTF and reduce MTTR.
Develop automation and operational tooling with Java/Python; champion cloud-native practices and observability to enhance service reliability.
Minimum Requirements
15+ years of professional experience in Site Reliability Engineering or related roles.
Must be located in or willing to work from Hyderabad, India.
Strong proficiency in software engineering with Java and/or Python for automation and tooling.
Experience with cloud-native platforms (Kubernetes/OpenShift), Infrastructure as Code, CI/CD, and advanced observability tools including AI-powered operational capabilities.
Ideal Candidate Profile
Seasoned SRE leader with expertise in defining and executing reliability strategies aligned to business objectives.
Deep technical background in software engineering automation coupled with cloud-native platform operations and observability.
Proven experience in technical leadership, mentoring, and cross-team collaboration to embed SRE best practices and drive operational excellence.
