Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessMid-level, popular SRE role in a metro with broad skills and a known employer.
Requires specific SRE/cloud experience but broadly transferable across software companies.
Explicit 5+ years and 2+ SRE years plus mandatory cloud, IaC, observability and Kubernetes requirements.
Job Description
Structured overview of role & requirementsAbout This Role
Lead reliability engineering initiatives to design and evolve scalable, highly available, and fault-tolerant infrastructure primarily on Azure cloud.
Define and enforce SLIs, SLOs, and error budgets; lead incident response and postmortems to improve system reliability and operational learning.
Develop and drive observability, automation, chaos engineering practices, mentor engineers, and embed SRE principles across global teams and product platforms.
Minimum Requirements
5+ years in software engineering with at least 2 years in Site Reliability Engineering, Platform Engineering, or a similar role.
Strong programming skills in Node.js, TypeScript, Go, Java, C#, or similar languages.
Experience with public cloud platforms (Azure preferred), Infrastructure as Code tools like Terraform/Pulumi, Kubernetes, and monitoring tools such as Prometheus and Grafana.
Applicant must be located in India; other locations may be declined.
Ideal Candidate Profile
Experienced in building and operating cloud-native, distributed systems with strong strategic ownership of reliability and operational excellence.
Skilled at collaborating across global, cross-disciplinary product and platform teams and influencing engineering practices at scale.
Proven mentor in SRE best practices and proactive in driving cultural and engineering process improvements globally.
