





Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Tier-1 brand, remote option, metro locations, and broad SRE/DevOps demand create high competition.
Deep cloud, Kubernetes, and SRE expertise required, limiting transferability across non-engineering industries.
Multiple explicit years and mandatory cloud, IaC, Kubernetes, and security requirements enforce strict screening.
Own and drive technical architecture and reliability of team-owned cloud infrastructure and services, ensuring scalability, cost-effectiveness, and security.
Lead incident response and establish operational standards including SLOs, logging, monitoring, and preventive measures to maintain service reliability.
Act as Directly Responsible Individual (DRI) for medium-to-large SRE projects, collaborating cross-functionally to scope, deliver, and mitigate project risks.
7-10 years experience in Site Reliability Engineering, DevOps, or Platform Engineering in production cloud environments.
5+ years hands-on experience with AWS services, Linux production environments, and Infrastructure-as-Code tools (Terraform, CDK, or CloudFormation).
3+ years in Kubernetes production operations, AWS Well-Architected Framework, cloud security practices, and PostgreSQL production management.
Work Experience Required: 7-10 years in relevant SRE or Platform Engineering roles.
Experienced leader capable of guiding technical architecture, incident management, and cross-team project delivery in cloud-native environments.
Strong AWS and Kubernetes operational expertise with solid programming skills in Python or Go for automation and tooling.
Knowledgeable in AI integration for system reliability improvements, including production use of LLMs, automation, and observability enhancements.