





Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Mid-level SRE in metro with broad cloud/devops requirements and recognizable employer drives high competition.
Core SRE skills transfer across industries, but deep networking and cloud knowledge require domain-specific experience.
Explicit 3–8 year requirement and mandatory Linux, cloud, networking, and scripting skills enforce high filters.
Own L2/L3 incident response, root cause analysis, and post-mortems for production issues to maintain system reliability and SLA/SLO adherence.
Automate operational tasks using scripting and tooling to reduce repetitive toil and improve efficiency.
Collaborate with development teams on deployment reliability, capacity planning, and participate in 24/7 on-call rotation maintaining runbooks.
3–8 years of experience in Site Reliability Engineering or related roles.
Hands-on experience with Linux (RHEL/Ubuntu) system management and performance tuning; working knowledge of Windows Server.
Practical experience with AWS, Azure, or GCP including compute, storage, IAM, networking, and Terraform or equivalent IaC tools.
Proficiency in Python and Bash scripting for automation and API interaction; deep understanding of TCP/IP networking and HTTP troubleshooting.
Experienced in managing critical production environments with operational accountability and incident escalation responsibilities.
Skilled in building observability dashboards and configuring alerts using Prometheus, Grafana, Datadog, or ELK for end-to-end issue tracing.
Comfortable with collaborative deployment processes and capacity planning in a fast-paced, 24/7 support environment.