





Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Specialized senior SRE in a metro with moderate brand and seniority, giving medium competition.
Core SRE and cloud skills are broadly transferable across industries, so background fit sensitivity is low.
Explicit 8+ years plus mandatory SRE, cloud, Kubernetes, and monitoring skills makes shortlisting strict.
Lead design and implementation of reliability solutions for complex distributed systems to enhance system uptime and performance.
Drive SRE best practices adoption including incident response, SLO/SLA definition, and error budget management across engineering teams.
Mentor junior engineers, conduct root cause analysis for production incidents, and automate operational tasks to improve efficiency.
Bachelor's or Master’s degree in Computer Science, Engineering, or related technical field.
8+ years of experience in software development, DevOps, or Site Reliability Engineering focused on system reliability and performance.
Proven expertise with at least one major cloud platform (AWS, Azure, or GCP) and strong programming skills in Python, Go, Java, or C++.
Experience with containerization (Docker, Kubernetes), monitoring tools (Prometheus, Grafana, ELK), and CI/CD pipelines including infrastructure as code (Terraform, Ansible).
Experienced in leading technical initiatives and mentoring engineers within reliability or Site Reliability Engineering domains.
Strong operational focus on designing highly available, scalable distributed systems in cloud environments.
Skilled in collaborating closely with development teams to embed reliability from inception and managing critical production incidents through structured processes.