





Tier-1 employer, mid-level SRE role, metro locations and broad cloud/observability requirements drive high competition.
Core DevOps/SRE skills are transferable, though pharma regulatory experience is a plus.
Explicit 5+ years and mandatory SRE, cloud, Terraform, and observability skills.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Own and improve enterprise AI Agent Platforms' reliability, performance, and operational excellence with a focus on cloud-native infrastructure.
Develop and maintain observability platforms, infrastructure as code (Terraform), and automation to enhance monitoring, alerting, incident response, and cost efficiency.
Collaborate with architecture, security, product teams, and stakeholders to drive platform roadmap, developer experience, and mentor engineers in best practices.
Bachelor’s degree in Computer Science, Data Science, Engineering, or related discipline.
5+ years of experience in Site Reliability, Operations, or Infrastructure engineering.
Hands-on experience with telemetry and observability tools (e.g., OpenTelemetry, Grafana), Azure and/or Google GCP services, Terraform, GitHub Actions, Kubernetes, Docker, and scripting (Bash, Python, Typescript).
Work Experience Required: 5+ years in relevant domains; Notice Period: Not explicitly mentioned in the JD.
Experienced in managing cloud-native platform operations with strong expertise in observability, automation, and incident response.
Comfortable working in dynamic and collaborative environments, balancing focused individual work with team-based problem-solving.
Skilled at collaborating with cross-functional teams including architecture, security, and product to influence platform strategy and support compliance frameworks.