





Tier-1 brand, popular data-engineer title, and Pune metro increase applicant competition.
Cloudera platform specificity reduces cross-industry portability despite transferable data engineering skills.
Explicit 6–9 years plus mandatory Cloudera, Spark, PySpark, Jenkins, and CI/CD skills enforce strict screening.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Administer and support Cloudera CDP/CDH platforms including HDFS, Hive, Spark, YARN, Hue, and CDE to ensure high availability and platform health.
Develop, deploy, and optimize PySpark and Python data processing solutions and write/optimize SQL queries for Hive and Impala.
Build and maintain CI/CD pipelines using Jenkins and GitHub/Bitbucket, integrating security tools like Checkmarx for automated quality and security checks.
6–9 years of experience in data engineering and DevOps, specifically with Cloudera Hadoop/CDP platforms.
Strong hands-on skills in HDFS, Spark/PySpark, Python, Hive, YARN, Hue, and CDE.
Experience with Jenkins, GitHub/Bitbucket CI/CD pipelines and integrating security tools such as Checkmarx.
Proficient in SQL query writing and performance tuning; strong Linux administration and scripting skills; located in Pune (Hybrid – Client Office).
Experienced in managing enterprise-scale Cloudera Big Data environments with responsibility for platform stability and upgrades.
Skilled in developing and optimizing automated data pipeline solutions using modern DevOps practices and tools.
Capable of implementing security controls and governance best practices in delivery pipelines to ensure compliance and code quality.