





Tier-1 employer, metro location, broad SRE skillset, and common SRE title drive high competition.
Data-focused SRE skills are transferable across industries but favor candidates with big-data platform experience.
Explicit 8+ years requirement plus mandatory SRE/DevOps and specific tooling increases shortlisting strictness.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Manage availability, latency, performance, and capacity planning of large-scale data platforms ensuring reliability and scalability.
Design and implement monitoring, automation tools, and incident response procedures to maintain and optimize data infrastructure.
Collaborate with engineering, vendors, and cross-functional teams to troubleshoot and resolve complex data platform issues and support data product implementation.
Bachelor’s degree in Computer Science, Software Engineering, or related field (or equivalent coursework/experience).
7-10 years of experience in Site Reliability Engineering, DevOps, or Data Operations engineering roles.
Proficiency with cloud platforms (AWS, GCP, Azure) and big data technologies such as Kafka, Hadoop, Spark, distributed storage systems (Cassandra, HDFS, AWS S3).
Experience with automation tools like Ansible, Terraform, Kubernetes, Docker, plus programming skills in Python, Go, Java or Scala.
Experienced in operating and scaling large-scale distributed data systems with a focus on reliability and performance.
Strong technical expertise to independently troubleshoot, debug, and optimize complex data pipelines and storage platforms.
Comfortable working collaboratively across multiple technical teams including data engineering, product, and operations.