





Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Broad observability and cloud skillset required at a mid-tier brand yields medium competition.
SRE and observability skills transfer across industries, though retail/payments domain experience is beneficial.
Multiple mandatory cloud, Kubernetes, observability, IaC, and language requirements enforce high shortlisting strictness.
Build and manage a unified observability platform providing end-to-end visibility across infrastructure, applications, Kubernetes, cloud services, and business transactions.
Design and implement enterprise observability solutions and standards including monitoring, logging, tracing, SLIs, SLOs, error budgets, and operational health metrics.
Integrate observability data with ServiceNow and automation platforms; collaborate across multiple teams to improve platform reliability and proactive incident response.
Experience with Site Reliability Engineering and Kubernetes (AKS/GKE).
Proficiency in Azure and Google Cloud Platform (GCP).
Skilled in observability tools such as Grafana, Datadog, Prometheus, OpenTelemetry or similar.
Work Experience Required: Not explicitly mentioned in the JD.
Expertise in enterprise-scale observability and telemetry across hybrid and multi-cloud environments.
Strong automation skills using Python, Go, or PowerShell and infrastructure as code (Terraform).
Experience working in cross-functional teams delivering reliability and operational analytics including AI-driven observability initiatives.