





Tier-1 brand, mid-level SRE role, metro location, and broad multi-cloud observability requirements increase competition.
Core SRE skills are broadly transferable, though industrial/OT preference increases domain specificity.
Explicit 3+ years plus mandatory observability, SRE tooling, GitOps and scripting requirements tighten filters.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Define and implement SLI/SLO frameworks for industrial software platforms including APM, SCADA, and OT-connected systems in multi-cloud environments (AWS and Azure).
Design and deploy observability pipelines using tools such as OpenTelemetry, Prometheus, Grafana, Loki, Jaeger and lead incident management including blameless postmortems with client engineering teams.
Build and scale SRE capabilities within client organizations through documentation, knowledge transfer, incident automation, chaos engineering, and cultural coaching.
3+ years of relevant work experience in SRE or related roles.
Proficiency in SLI/SLO design, error budget management, observability toolchains (Prometheus, Grafana, Loki, Jaeger, OpenTelemetry) across AWS and Azure.
Experience in incident management, on-call process design, postmortem facilitation, and automation/scripting with Python, Go, or Bash.
Bachelor's or Master's degree (field not specified).
Experienced in operating within industrial or OT/IT environments such as energy, manufacturing, or connected products.
Able to lead strategic transformation from reactive operations to automated reliability through toolchain integration and cultural change.
Skilled in GitOps and CI/CD practices including tools like ArgoCD, GitHub Actions, and Flux to enable scalable, automated SRE workflows.