





Senior, metro role but specialized observability skills and Kubernetes reduce general applicant density.
Observability and platform skills transfer across tech firms but remain domain-specific, so moderate portability.
Mandatory 9+ years plus specific observability, Kubernetes, cloud, and Go/Python/Terraform requirements increase filter strictness.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Own the design, build, and operation of Mercari's observability platform covering metrics, logs, traces, and alerting at scale for 500+ microservices.
Drive measurable improvements in Mean Time to Detect (MTTD) and Mean Time to Mitigate (MTTM) through AI-powered automated incident detection, alert correlation, and response.
Lead technical direction, mentor team members, define observability standards, and develop self-service tooling to enhance developer experience and platform reliability.
9+ years of experience building, operating, and maintaining scalable production systems.
Strong expertise with observability and monitoring platforms like Datadog, Prometheus, Grafana, in production environments.
Proficiency in Go or Python, Kubernetes, cloud platforms (GCP and/or AWS), and Infrastructure as Code (Terraform).
Deep understanding of metrics, logging, distributed tracing, alerting systems, SLIs/SLOs, and developer tooling for observability.
Experienced in scaling observability for large distributed systems with 500+ microservices and reducing noise in alerting systems.
Track record of improving operational metrics (MTTD and MTTM) using AI/automation in observability.
Strong leadership capabilities to shape technical direction, mentor engineers, and cultivate a strong engineering culture in platform teams.