





Tier-1 brand, hybrid metro setting, and broad in-demand SRE/ML infra skillset.
Role requires specialized ML/data-platform and SRE expertise, limiting cross-industry transferability.
Mandatory cloud, IaC, Kubernetes, Hadoop/Spark, and programming requirements enforce strict technical screening.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Lead design, build, and optimization of cloud and data infrastructure ensuring high availability, reliability, and scalability of big data and AI/ML platforms.
Provide technical leadership by defining and executing the team's technical roadmap, collaborating across teams to deliver secure, scalable SaaS solutions at multi-region scale.
Mentor teams and troubleshoot production issues, driving operational excellence through monitoring, fault analysis, and continuous improvement.
Strong cloud experience, preferably AWS, with Infrastructure as Code expertise (Terraform, Kubernetes/EKS).
Expertise with Hadoop ecosystem (Spark, Hive, HDFS, Gobblin), Airflow, and AWS big data services (EMR, SageMaker) for AI/ML infrastructure at scale.
Programming skills in Python, Go, or equivalent languages.
Work Experience Required: Not explicitly mentioned in the JD.
Experienced in architecting and scaling production AI/ML infrastructure with ownership and accountability.
Proficient in operationalizing large-scale data platforms integrating cloud-native and big data technologies.
Skilled at collaborating cross-functionally and leading technical strategy in a fast-paced, multi-region SaaS environment.