Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessMetro-based generalist data engineer role with broad skills and mid-level experience driving high competition.
Core data engineering skills are transferable across industries, so background fit sensitivity is low.
Multiple explicit years requirements and mandatory tech stack (PySpark, Snowflake, Databricks, AWS) increase shortlisting strictness.
Job Description
Structured overview of role & requirementsAbout This Role
Design and implement performant data ingestion pipelines from multiple sources ensuring data quality and consistency.
Develop scalable and reusable frameworks for data ingestion and integrate end-to-end data pipelines to target repositories.
Evaluate and demonstrate key technology components to stakeholders; work with event-based/streaming technologies and CI/CD tools for data processing.
Minimum Requirements
Bachelor’s degree in computer science, engineering, or related field mandatory.
At least 5 years hands-on experience with PySpark, Python, Spark, Scala, and data validation/delivery.
Minimum 5 years’ experience with Snowflake or Databricks Cloud Data Warehouse and handling multiple data file formats (XML, JSON, Avro, Parquet).
Work Experience Required: 5+ years relevant data engineering experience; Must be able to work 2:00 PM to 11:00 PM IST.
Ideal Candidate Profile
Experienced in Big Data models and ETL using Python/Spark with knowledge of structured, semi-structured, and unstructured data.
Proficient with modern schedulers (Airflow, Matillion), distributed systems (Hadoop, Spark, Kafka), and CI/CD tools (Jenkins, Terraform, Docker, AWS, Kubernetes).
Familiarity with cloud technologies on AWS (EC2, S3, Lambda), strong SQL programming skills, and understanding of data warehousing concepts.
