





Strong brand, remote/hybrid role, mid-level SRE, metro location, and broad cloud/DevOps skill requirements.
Core cloud, IaC, and SRE skills are transferable, but AI-platform specialization increases domain specificity.
Explicit 5+ years, mandatory cloud/infra, IaC and Kubernetes requirements make screening highly selective.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Own and drive long-term infrastructure strategy for AI Platform, ensuring scalability, reliability, and reduced operational burden.
Manage operational posture of cloud-native services including containerized workloads, serverless functions, and managed ML infrastructure, incorporating observability and ensuring SLA compliance.
Design, build, and maintain CI/CD pipelines; automate toil such as infrastructure provisioning, security remediation, and deployment operations; mentor junior team members.
5+ years of experience in cloud and infrastructure engineering roles.
Strong proficiency in Python and shell scripting; experience with infrastructure-as-code tools like CloudFormation, Terraform, ARM, or SAM.
Hands-on experience with cloud infrastructure, containerization (Docker, Kubernetes), and CI/CD tooling.
Work Experience Required: 5+ years
Experienced in designing and scaling cloud-native infrastructure for ML/AI platforms with emphasis on reliability and automation.
Operates effectively across engineering, security, and product teams to translate requirements into engineering outcomes.
Skilled in building automation and deployment pipelines to streamline DevOps workflows and improve operational efficiency.