Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessMetro location, broad SRE skillset, and popular SRE title create high applicant competition.
Core cloud and SRE skills transfer across industries, so background sensitivity is low.
Technical requirements (GCP, Kubernetes, Terraform, CI/CD) but no explicit years requirement makes shortlisting moderately strict.
Job Description
Structured overview of role & requirementsAbout This Role
Provide on-call support for Reservations suite, handle alerts, troubleshoot and lead incident investigations impacting customer experience.
Build and maintain monitoring and alerting solutions in GCP and related observability tools; improve service reliability through SRE practices including automation and toil reduction.
Own service reliability risks, infrastructure management tasks (PCI audits, capacity monitoring, cost optimization), support infrastructure deployments, CI/CD pipelines, and mentor team members.
Minimum Requirements
Experience with Google Cloud Platform and SRE principles, including CI/CD, container orchestration (Docker, Kubernetes), Linux/UNIX, and scripting (Terraform, shell).
Familiarity with monitoring tools (AppDynamics, Google Cloud Ops, DataDog, Prometheus, Grafana, Elasticsearch).
Strong troubleshooting, debugging, and change management expertise, good English communication skills.
Work Experience Required: Not explicitly mentioned in the JD.
Ideal Candidate Profile
Comfortable operating in a hybrid global team supporting mission-critical travel reservation systems with high availability requirements.
Experienced in proactive reliability engineering with strong automation and incident leadership capabilities.
Able to mentor peers and enforce best practices related to production readiness and service reliability.
