





Metro Bangalore, popular platform/SRE role, moderate brand and mid-level seniority.
Cloud, Kubernetes, and SRE skills are highly transferable across industries.
Multiple mandatory cloud, Kubernetes, programming, and observability requirements create stringent shortlisting filters.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Design, build, and operate highly reliable, scalable infrastructure for Guidewire's multi-tenant SaaS platform, focusing on automation, reliability, scalability, and operability.
Develop internal tools, services, and frameworks to improve efficiency, and participate in 24x7 follow-the-sun on-call rotation supporting critical production systems.
Lead observability, incident response, and security efforts including maintaining metrics, dashboards, Service Level Objectives, and secure access patterns (SSO, SAML, OAuth).
Strong programming skills in Python or Go; Java/Spring Boot is a plus.
Deep experience with AWS, Kubernetes (EKS), Docker, Helm, and Infrastructure as Code tools like Terraform or Terragrunt.
Experience in production support of large-scale SaaS or distributed systems, including observability and incident management.
Bachelor's degree in Computer Science or related field, or equivalent experience.
Experienced Site Reliability Engineer with expertise operating global, multi-tenant SaaS platforms at scale using Kubernetes and AWS.
Proven ability to design and automate scalable infrastructure with strong systems-thinking and troubleshooting skills.
Skilled collaborator able to influence development teams, lead incident response, and mentor peers in reliability engineering and automation.