Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Own reliability, availability, and performance of microservices and production workloads on GCP.
Design and improve resilient cloud infrastructure using Cloud Run, Kubernetes, with a focus on CI/CD, incident response, and production readiness.
Drive automation to reduce operational toil, enhance observability, cloud security, cost management, and support architecture and reliability reviews.
Minimum Requirements
5+ years experience in Site Reliability Engineering or related DevOps roles with production ownership.
Strong production experience with Google Cloud Platform, including Cloud Run and Kubernetes.
Hands-on with infrastructure as code tools such as Terraform and Terragrunt.
Experience developing and managing CI/CD pipelines with tools like GitHub Actions.
Ideal Candidate Profile
Experienced in operating distributed microservices with a strong understanding of observability and reliability engineering principles.
Skilled in automation scripting (Python or Java) and cloud security practices focused on IAM, secrets management, and workload hardening.
Comfortable leading incident response activities and collaborating with dev teams to improve scalability, fault tolerance, and production readiness.
