





Strong employer brand, metro location, mid-level experience, and generalist SRE skillset increase applicant competition.
Core SRE and monitoring skills are fairly transferable across industries despite hotel domain context.
Explicit 3+ years plus mandatory production, monitoring, SQL and scripting skills create moderately strict filters.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Lead investigations into high-impact incidents and recurring operational issues affecting hotel partner experiences, performing root cause analysis and driving corrective actions.
Analyze system behavior, logs, metrics to resolve complex production issues in distributed, high-availability environments and partner with multifunctional teams to implement sustainable reliability improvements.
Take ownership during live incidents including leading SRE bridge calls, support weekend on-call rotations when activated, and act as escalation point and mentor for junior team members.
Bachelor’s or master’s degree in a technical discipline or equivalent related professional experience.
Minimum 3+ years experience supporting production or operational environments with demonstrated leadership in investigations and operational improvements.
Experience with monitoring and observability tools like Splunk, Datadog, Kibana, and proficiency in SQL for investigation and data analysis.
Familiarity with API architectures (such as GraphQL), cloud and infrastructure concepts, testing principles, and at least one scripting or programming language.
Comfortable operating in high-availability, distributed systems environments with strong analytical and troubleshooting skills.
Experienced in cross-functional collaboration with engineering, product, and business teams to implement operational improvements with measurable impact.
Able to lead incident response and mentor junior staff, demonstrating ownership and a systems-level problem-solving approach.