Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Ingest and transform data from multiple sources into a Data Lake using Java-based tools like Apache Spark and Spring Batch.
Develop and modernize ETL pipelines, including writing Spark (Java API) transformation logic and stored procedures for relational databases.
Handle both batch and real-time data processing, storing output in HDFS or cloud storage (AWS S3, GCP Storage, or on-prem).
Minimum Requirements
5+ years of relevant software engineering experience.
Proficient in Java programming, including Spark code development and SQL query writing.
Experience with Java frameworks for data processing and integration such as Spring, Apache Spark, JDBC.
BE/BTech degree required.
Ideal Candidate Profile
Experienced in data warehousing concepts and data modeling techniques like star schema.
Familiar with ETL workflows, including batch and real-time processing and transformation.
Understands advanced data management like Delta Lake transactional storage and Slowly Changing Dimensions (SCD).
