





Medium competition: popular SRE role and Bangalore metro increase applicant density.
SRE and cloud skills are broadly transferable, though cloud-native SaaS context adds moderate domain bias.
High strictness: many mandatory technical skills and a senior 8+ years requirement.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Own the reliability and operational health of large-scale Kubernetes-based cloud infrastructure supporting critical business applications.
Design, build, and maintain observability dashboards, alerting systems, and automation to improve system reliability and reduce manual operational work.
Lead incident management including root-cause analysis and remediation, and enforce SRE best practices and service level objectives across engineering teams.
8-10+ years of relevant hands-on Site Reliability Engineering or infrastructure engineering experience.
Strong hands-on expertise with Kubernetes and cloud platforms (AWS, Azure, or GCP) including networking fundamentals.
Practical experience defining and monitoring SLIs, SLOs, and managing error budgets in production systems.
Experience with observability tools (e.g., Prometheus, Grafana, Datadog), CI/CD systems (preferably GitHub Actions), and handling live production incidents under pressure.
Experienced in operating and improving reliability of large distributed systems with complex failure modes and significant scale.
Comfortable owning diverse operational problem areas and making decisions balancing immediate fixes and long-term scalable solutions.
Skilled in AI-assisted engineering tools for automation and root cause analysis applied in live production environments.