





Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Mid-level PySpark data engineer in metro with common skillset and typical experience, making competition high.
PySpark and CDP skills transfer across industries but favor big-data platform roles, so medium sensitivity.
Explicit 5+ years preferred plus mandatory PySpark, CDP, and Hadoop skills increases shortlisting strictness to high.
Design, develop, and maintain scalable data pipelines using PySpark and Cloudera Data Platform (CDP).
Build, optimize, and manage data ingestion, transformation processes, and data quality assurance across big data platforms.
Collaborate with cross-functional teams and monitor data workflows for performance and reliability issues.
Strong experience with PySpark and Cloudera Data Platform (CDP).
Hands-on experience with Hadoop, Hive, HDFS, Spark, and ETL processes.
Good understanding of data warehousing and big data architecture, including cloud and distributed data processing environments.
Work Experience Required: 5+ years experience preferred, Location: Chennai
Experienced data engineer with proven expertise in big data ecosystems and cloud-native environments.
Ability to manage large-scale datasets and complex data workflows under operational constraints.
Demonstrated skills in optimizing data pipelines and ensuring data integrity and availability.