IN_Senior Associate_Databricks Data Engineer_GCC_Advisory_Gurgaon
PwC IndiaMatch Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Design, build, and maintain scalable, high-performance data pipelines using Databricks Lakehouse Platform and Apache Spark technologies.
Ingest, transform, and manage large-scale structured and semi-structured datasets ensuring data quality, lineage, governance, and compliance.
Collaborate with analytics, AI/ML, and business teams to support data consumption, optimize Spark jobs for performance and cost, and provide production support and mentorship.
Minimum Requirements
4+ years of experience as a Data Engineer with strong Databricks expertise.
Hands-on skills with Apache Spark, PySpark, Spark SQL, Delta Lake, and ETL/ELT data warehousing concepts.
Bachelor's or Master's degree in Computer Science, Engineering, or a related field with 60% or above.
Experience with at least one cloud platform (Azure, AWS, or GCP).
Ideal Candidate Profile
Experienced in developing and optimizing data pipelines on Databricks with knowledge of Lakehouse architecture and advanced SQL.
Familiarity with Agile environments and collaborative work with data scientists, analysts, and architects on AI/ML workloads.
Possesses certifications or exposure to Databricks tools like Unity Catalog, Auto Loader, CI/CD for data pipelines, and Python for automation.
