Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Design, build, and run reliable, scalable, distributed data pipelines and platforms using Databricks and AWS.
Model data for analysis ensuring accuracy, governance, and usability while meeting privacy and compliance requirements.
Mentor engineers, act as an informal technical lead within the squad, and maintain platform reliability through monitoring, alerting, and business-hours support.
Minimum Requirements
Strong hands-on experience with Databricks, PySpark, Spark SQL, Delta Lake, Unity Catalog, and AWS data/serverless services like Lambda, Glue, S3, Redshift, Kinesis.
Proficiency in Python, advanced SQL, data modeling, infrastructure as code, and CI/CD pipeline management.
Experience mentoring engineers and influencing cross-functional teams technically.
Degree in computer science or related field, or equivalent knowledge and experience. Work Experience Required: Not explicitly mentioned in the JD.
Ideal Candidate Profile
Experienced in complex environments managing distributed data pipelines and lakehouse technologies at scale.
Capable of technical leadership through mentoring and raising engineering quality without formal management authority.
Comfortable collaborating across business units to deliver dependable data products and committed to continuous improvement mindset.
