IN-Sr Associate_Databricks Data Engineer_GCC_Advisory_Bangalore
PwC IndiaMatch Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Design, build, and maintain scalable, high-performance data pipelines on the Databricks Lakehouse Platform using PySpark, Spark SQL, and Delta Lake.
Collaborate with analytics, AI/ML, and business teams to support analytics, reporting, and downstream data consumption ensuring data quality, reliability, and governance.
Provide production support including troubleshooting data pipeline issues and mentoring junior engineers while documenting technical designs and operational runbooks.
Minimum Requirements
4-7 years of experience as a Data Engineer with strong hands-on expertise in Databricks, Apache Spark, PySpark, and Spark SQL.
Bachelor's or Master's degree in Computer Science, Engineering, or related field with 60% or above (B.Tech/B.E. preferred).
Strong knowledge of Delta Lake and Lakehouse architecture, advanced SQL skills, and experience with cloud platforms (Azure / AWS / GCP).
Not explicitly mentioned: notice period or location constraints explicitly stated.
Ideal Candidate Profile
Experienced in implementing batch and incremental data processing using Delta Lake and multi-hop architecture optimizing Spark jobs for performance and cost.
Comfortable working in Agile teams with exposure to data governance tools such as Unity Catalog, streaming frameworks like Auto Loader, and CI/CD for data pipelines.
Holds or pursues relevant Databricks certifications and proficient in scripting languages such as Python for data engineering and automation.
