





Mid-level SRE title, metro location, and broad skillset create high applicant competition.
Core SRE tooling and cloud skills transfer across industries, but platform-specific knowledge raises sensitivity.
Explicit 5+ years and 2+ SRE years plus mandatory cloud/IaC/Kubernetes tooling increases strictness.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Lead and mentor SRE team members while influencing product teams to build reliable, scalable, and observable systems from design to delivery.
Define and maintain SLIs, SLOs, error budgets, and drive incident management processes including blameless postmortems and operational excellence.
Develop and automate infrastructure as code, monitoring, chaos engineering, and release validation pipelines to support large-scale, fault-tolerant platform operations across public cloud environments, initially Azure.
5+ years of software engineering experience with at least 2 years in Site Reliability, Platform Engineering, or related roles.
Experience with public cloud platforms, preferably Azure.
Proficient in programming languages such as JavaScript, Node.js, Typescript, Go, Java, or C# and experience with monitoring and observability tools like Prometheus, Grafana, OpenTelemetry.
Experience with Infrastructure as Code (Terraform/Pulumi) and container orchestration platforms such as Kubernetes.
Technical leader comfortable operating across SRE, platform engineering, and product collaboration to embed reliability practices organization-wide.
Experienced with cloud-native distributed system design and skilled in building automated tooling for observability, incident response, and self-healing.
Capable of leading strategic initiatives, mentoring senior engineers, and scaling SRE best practices across global teams in a complex, multi-cloud environment.