IN_Manager_Databricks Data Engineer_GCC_Advisory_Bangalore
PwC IndiaMatch Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Design, build, and maintain scalable, high-performance data pipelines on Databricks Lakehouse Platform using Apache Spark (PySpark/Spark SQL) and Delta Lake.
Collaborate with analytics, AI/ML, and business teams to support data-driven insights and downstream data consumption while ensuring data quality, governance, and security compliance.
Provide production support, troubleshoot pipeline issues, optimize Spark jobs for performance/cost, document technical designs, and mentor junior engineers.
Minimum Requirements
7+ years of experience as a Data Engineer with strong expertise in Databricks, Apache Spark, PySpark, and Spark SQL.
Bachelor’s or Master’s degree in Computer Science, Engineering, or related field with 60% or above.
Experience with at least one cloud platform (Azure, AWS, or GCP) and strong knowledge of Delta Lake and Lakehouse architecture.
Advanced SQL skills and understanding of distributed computing concepts.
Ideal Candidate Profile
Experienced in managing end-to-end data pipeline lifecycle in Agile environments, including data ingestion, transformation, and pipeline optimization on Databricks.
Familiar with enterprise security, access control, compliance standards, and production-level support for data engineering workloads.
Capable of mentoring junior engineers and contributing to best practices including documentation and code optimization.
