





Mid-level SRE, broad skillset required, and recognizable global employer increase candidate competition.
Core SRE skills are transferable across industries though retail/payments command center experience is preferred.
Explicit 2+ years, mission-critical SRE experience, and specific cloud and monitoring tooling raise filters.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Ensure availability, reliability, and performance of business-critical applications and customer-facing services in a 24x7 global operational environment.
Lead proactive monitoring, incident response, problem management, and operational readiness reviews to minimize service disruptions and improve resiliency.
Develop and maintain automation and monitoring solutions across hybrid cloud platforms (Azure, GCP, AWS) to improve operational efficiency and service reliability.
Minimum 2 years experience in Site Reliability Engineering, Production Support, Systems Engineering, DevOps, NOC, or Command Center Operations.
Experience supporting mission-critical production environments with strong understanding of incident, problem, change, and service level management.
Proficiency with monitoring tools such as AppDynamics, Datadog, Dynatrace, Splunk, Azure Monitor, and ServiceNow Event Management.
Technical experience with Windows Server, Linux/Unix, and cloud platforms (Azure, GCP, AWS).
Experienced working in global command center or 24x7 operational support environments managing on-call and major incident rotations.
Skilled in reliability engineering practices including SLI/SLO/SLA management, automation, and resilience improvement initiatives.
Strong ability to collaborate across application, infrastructure, cloud, and network teams while driving continuous operational improvements.