





Medium - metro mid-level SRE with broad cloud, Kubernetes and automation requirements.
Medium - core SRE/cloud skills transferable, but reliability practices and domain knowledge increase specificity.
High - explicit 5+ years and mandatory cloud, Kubernetes, IaC, and SRE domain skills.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Design, deploy, and maintain scalable, secure applications and infrastructure in cloud or hybrid environments.
Lead incident response through on-call rotations, root cause analysis, and implement preventive and self-healing measures to meet RTO and RPO targets.
Automate operational tasks and collaborate with engineering and product teams to improve service availability, reliability, and operational processes.
Bachelor’s degree in Computer Science, Systems, Electrical Engineering, or relevant discipline.
5+ years experience in Site Reliability Engineering, DevOps, or related role managing large-scale solutions.
Proficiency in scripting languages (PowerShell, Python, Go, Bash) and hands-on experience with cloud platforms (AWS, GCP, Azure) and container orchestration (Docker, Kubernetes).
Experience with monitoring, alerting, observability tools, CI/CD pipelines, Agile methodologies, and incident response processes.
Experienced in managing large-scale cloud or hybrid infrastructure with strong automation skills using Infrastructure-as-Code tools like Terraform.
Comfortable working in high-pressure, production-critical environments with a focus on operational excellence and reliability.
Able to collaborate across engineering, product, and senior stakeholders to influence technical decisions and drive organizational reliability improvements.