Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessGeneralist Data Engineer title, metro location, and broad skill requirements drive high candidate competition.
Strong platform- and migration-specific requirements (HDFS, Snowflake, Iceberg, temporal modeling) reduce cross-industry portability.
Explicit 8+ years plus many mandatory platform-specific skills and tools enforce strict shortlisting.
Job Description
Structured overview of role & requirementsAbout This Role
Own the end-to-end migration of datastore from on-prem DataLake (Hadoop/HDFS) to AWS-hosted Lakehouse using tools and pipelines.
Manage CI/CD pipelines related to data migration including investigation and root cause analysis of pipeline logs.
Translate and optimize legacy SQL and Spark queries for compatibility with Snowflake and Apache Iceberg platforms.
Minimum Requirements
8+ years of experience in Data Engineering, with 3–5 years of hands-on coding experience in a team environment.
Bachelor's or Master's degree in Computer Science, Applied Mathematics, Engineering, or related quantitative field.
Proficiency in Kafka, ANSI SQL, FTP, Apache Spark, Python/Java, RESTful APIs, and experience with CI/CD pipelines.
Experience with on-prem Hadoop ecosystem (HDFS), AWS S3 data staging, Snowflake, Apache Iceberg, and handling schema evolution management.
Ideal Candidate Profile
Experienced in executing complex legacy to cloud datastore migration projects involving data reconciliation and temporal data modeling (e.g., SCD Type 2).
Skilled in troubleshooting and root cause analysis of data migration tools and pipelines with ability to coordinate with multiple teams.
Comfortable working in high-visibility projects requiring deep expertise in both on-prem Hadoop and AWS cloud data platform ecosystems.
