





Tier-1 employer, metro location, and broad data/AI skillset increase applicant competition.
Medium because data-engineering skills transfer across industries but require platform-specific expertise.
Medium due to firm requirements for deep Spark/Python/data platform skills despite no explicit years filter.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Design and build scalable ETL/ELT pipelines using Python and PySpark to handle large-scale data ingestion and transformation.
Integrate Generative AI and Machine Learning models into data workflows to enhance platform capabilities and solve complex technical challenges.
Implement CI/CD pipelines, performance tuning, and enforce data quality standards to ensure reliability and efficiency of data processing systems.
Strong proficiency in Python and Apache Spark, including Spark Core and Spark SQL, applied in production environments.
Advanced SQL skills with experience in complex query writing and performance optimization.
Demonstrated experience designing and maintaining large-scale ETL/ELT pipelines.
Work Experience Required: Not explicitly mentioned in the JD.
Experience working in multidisciplinary engineering teams with a focus on delivering scalable data solutions.
Technical background with hands-on exposure to DevOps practices, CI/CD automation, and Linux environments.
Knowledge or practical experience with AI technologies including Generative AI, Machine Learning, and Agentic AI workflows to drive innovation in data platforms.