





Tier-1 brand, hybrid remote, mid-level SRE role in metro with broad cloud/Kubernetes skills.
Core cloud, Kubernetes, and automation skills transfer across industries but SRE experience has moderate domain specificity.
Explicit 2–3 year requirement plus mandatory cloud, Kubernetes, IaC, CI/CD skills and graveyard-shift readiness.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Own operational reliability and health monitoring of large-scale enterprise software production environments.
Develop and maintain software and systems for platform infrastructure management, automation, and performance optimization.
Lead incident management and capacity planning, partner with development teams to enhance system reliability and deployment processes.
2-3 years experience in site reliability, systems engineering, or automation-focused roles.
Bachelor’s degree in Computer Science, Engineering, or related field, or equivalent experience.
Proficiency in at least one programming language (Go, Python, .Net C#) and scripting (Bash, PowerShell).
Experience with cloud platforms (AWS), infrastructure as code (CloudFormation, Terraform), containerization (Docker, Kubernetes), CI/CD tools, and monitoring/observability tools.
Experienced with managing large distributed software systems in cloud environments with focus on reliability and automation.
Skilled in incident response and driving cross-functional collaboration during outages.
Comfortable working night shifts and balancing rapid feature delivery with system stability.