





Mid-level, metro SRE role at a recognizable brand with broad, popular skill requirements increases competition.
SRE skills are transferable across industries but specialist reliability and networking expertise moderately constrain fit.
Explicit 3–8 year requirement plus mandatory SRE/cloud/tooling skills increases filter strictness.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Own L2 incident response, root cause analysis, and post-mortems for production issues to maintain system reliability.
Monitor system health to ensure SLA/SLO compliance and automate operational tasks to reduce repetitive toil.
Participate in on-call rotations and collaborate with development teams on deployment reliability and capacity planning.
Experience Required: 3–8 years in Site Reliability Engineering or related roles.
Hands-on experience with Linux (RHEL/Ubuntu), and working knowledge of Windows Server.
Practical experience with AWS, Azure, or GCP cloud platforms including compute, storage, IAM, networking, and managed services.
Proficiency in Python and Bash scripting, strong command over TCP/IP network troubleshooting, and experience with observability tools like Prometheus, Grafana, Datadog, or ELK.
Experienced in managing production infrastructure with incident response accountability and on-call responsibilities.
Comfortable working 24x7 shift rotations and collaborating across teams for deployment and capacity planning.
Strong operational focus on automation through scripting and tooling to minimize manual toil and improve system performance.