Lead Software Engineer - Pyspark, AWS, Python
JPMorgan Chase & Co.Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Lead design, development, and maintenance of scalable cloud-based data processing pipelines and infrastructure, ensuring industry best practices and data governance compliance.
Architect and optimize large-scale data models and infrastructure for efficient analytics and high data quality.
Drive data strategy execution, process re-engineering, and adoption of AI-assisted engineering tools to improve code quality, delivery speed, and operational outcomes.
Minimum Requirements
5+ years of applied experience with formal training or certification in software engineering concepts.
Expertise in distributed data processing frameworks (Spark), cloud data lake platforms (AWS, Databricks), and scheduling/orchestration tools (Airflow or similar).
Proficiency in Python, SQL, and at least one additional programming language such as Java or Scala.
Hands-on experience with microservices, serverless computing, distributed cluster tools (Docker, Kubernetes), data modeling techniques, CI/CD, test-driven/behavior-driven development, and streaming platforms like Kafka.
Ideal Candidate Profile
Experienced technical leader capable of architecting reusable, future-ready data engineering solutions across diverse organizational use cases.
Demonstrated ability to lead cross-functional teams, organize design workshops, and promote a culture of excellence and innovation in data engineering.
Proficient in responsible AI-assisted development workflows, balancing automation with security, compliance, and team coaching.
