





Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Tier-1 brand, metro locations, and broad DevOps skillset increase candidate competition.
Requires cloud, networking, and SRE skills, moderately limiting cross-industry transferability.
Explicit 8+ years requirement and specific cloud, networking, and automation skills enforce strict shortlisting.
Respond promptly to and resolve critical service outages impacting consumers’ experience.
Develop and automate monitoring, alerting, and capacity models to prevent service interruptions and optimize system performance.
Investigate root causes of outages, implement corrective actions, and drive architecture improvements to enhance overall availability and reliability.
Bachelor's Degree in Computer Science or STEM fields.
For USA-based roles: minimum 8 years of relevant experience; For roles outside USA: advanced experience required (not explicitly quantified).
Strong hands-on expertise in scripting/programming languages such as Ruby, Python, Go, Java, Node.js, or .NET.
Experience with cloud infrastructure (AWS/Azure), automated configuration management tools (Terraform, Chef, Puppet, Ansible, Salt), network protocols (TCP/IP, SNMP, etc.), and monitoring tools (Datadog, Sensu, Grafana, Splunk).
Technically strong with a focus on automation to ensure high uptime and availability in complex distributed systems.
Experienced in cloud environments and network management to proactively identify and remediate risks before outages occur.
Skilled in developing capacity planning and operational health checks aligned with consumer experience metrics to drive improvements.