Match Score
Against your primary resumeLogin to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Protocol Intelligence
Data-driven signals on your job's competitivenessLog in to see why each signal reads the way it does.
Job Description
Structured overview of role & requirementsAbout This Role
Design, develop, and maintain scalable data processing and analytics services using big data technologies such as Spark, Trino, Airflow, and Kafka.
Operate and improve resilient distributed systems managing thousands of compute nodes across multiple data centers.
Own full service lifecycle including live-site reliability, feature development, technical debt reduction, and participate in on-call rotations for critical service availability.
Minimum Requirements
Proficiency in cloud environments (AWS, GCP, or Azure) including containerization technologies (Docker, Kubernetes) and infrastructure-as-code tools (Terraform, Ansible).
Hands-on experience with big data technologies like Hadoop, Spark, Trino or similar SQL query engines, Airflow, and Kafka.
Strong programming skills in Python, Java, Scala, or equivalent languages relevant to distributed systems.
Work Experience Required: Not explicitly mentioned in the JD.
Ideal Candidate Profile
Experienced with distributed systems principles including data partitioning, fault tolerance, and performance tuning for large-scale infrastructure.
Skilled at troubleshooting complex system issues and optimizing for efficiency and scalability.
Comfortable with operational responsibilities such as live-site support and maintaining availability across multi-data center environments.
