





Specialized SRE/observability skills but popular SRE hiring increases applicant density.
Highly domain-specific observability and SRE expertise required, limiting cross-industry transferability.
Explicit 8+ years plus mandatory Grafana Cloud, SLIs/SLOs, IaC, cloud and Kubernetes skills.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Own end-to-end design, administration, and optimization of Grafana Cloud observability platform.
Lead migration and consolidation of observability tools onto Grafana Cloud while supporting existing Datadog environments.
Define and implement observability standards, SLIs, SLOs, SLAs, error budgets, and automate platform provisioning to enhance service reliability and reduce alert noise.
8+ years experience in SRE, observability, or monitoring engineering.
Bachelor’s degree in computer science, information systems, or related field.
Expert-level experience with Grafana Cloud; hands-on experience with Datadog, Dynatrace, ELK, or Logz.io.
Strong understanding of SLIs, SLOs, SLAs, error budgets, observability automation using Infrastructure as Code, APIs, CI/CD, and cloud platforms (Azure, AWS, GCP).
Experienced in leading observability platform architecture and global transformation efforts.
Skilled in implementing reliability engineering practices (SLIs/SLOs) and observability automation to improve incident response.
Comfortable collaborating cross-functionally with application, SRE, and DevOps teams in a large, global, multi-industrial environment.