Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Own and contribute to reliability architecture and standards for distributed, microservices, and event-driven cloud platforms.
Design and implement highly available systems with multi-region deployments, disaster recovery, and failover capabilities.
Drive operational excellence via SLIs/SLOs, incident management, RCA, observability tooling, and continuous improvement in cloud infrastructure automation and platform resilience.
Minimum Requirements
Minimum 5 years experience in SRE, DevOps, Platform Engineering, Cloud Engineering or related roles.
Proven experience designing and operating cloud-based SaaS distributed systems with cloud-agnostic architecture understanding.
Hands-on expertise with Kubernetes, Terraform or Bicep, CI/CD pipelines, infrastructure automation, public cloud platforms.
Experience with observability tools (Prometheus, Grafana, OpenTelemetry) and messaging technologies (Azure Service Bus, Kafka, RabbitMQ).
Ideal Candidate Profile
Experienced in building resilient, multi-region cloud platforms balancing architectural vision and operational delivery.
Comfortable working with microservices, event-driven architectures, and messaging platforms across multiple programming languages (C#, Python, TypeScript, Go).
Able to mentor teams, contribute technical decisions, and maintain composure during critical incidents while fostering continuous improvement practices.
