





High due to Tier-1 brand, metro Hyderabad location, and sought-after ML/data engineering skillset.
Medium because core ML and data engineering skills transfer, though life-sciences experience is preferred.
High due to explicit 7-11 years requirement and mandatory ML, PySpark, AWS, and data platform skills.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Design, develop, and maintain scalable data pipelines and platforms supporting AI/ML applications and advanced analytics.
Ensure pipeline reliability, performance, and governance through automation, monitoring, logging, and documentation.
Collaborate with data scientists, analysts, and business teams to translate requirements into production-grade, observable data solutions within a matrix environment.
Bachelor's degree in Information Technology, Computer Science or related technology stream.
7-11 years of experience in developing data pipelines and data infrastructure, ideally in drug development or life sciences.
Strong programming skills in Python (including PySpark) and SQL; experience integrating ML models into production.
Hyderabad location required with hybrid work model (3 days onsite, 2 days remote); candidate must reside within commuting distance.
Experienced in operating within matrixed global organizations enabling multi-stakeholder collaboration and project delivery.
Demonstrated expertise in cloud technologies (AWS S3, Redshift, Snowflake) and data engineering best practices focused on scalable and resilient production systems.
Skilled in experiment design and validation techniques to ensure data and pipeline quality, aligned to business goals in life sciences or biopharma domain.