





Remote mid-level data role in a metro market with generalist requirements increases competition.
Data engineering and cloud infrastructure skills are moderately transferable across industries.
Mandatory 5+ years plus required GCP, Terraform, Docker, and scripting enforces strict shortlisting.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Own end-to-end data collection and ingestion pipelines to support AI model training at petabyte scale.
Operate and extend cloud infrastructure (GCP, managed with Terraform) for audio data ingestion.
Collaborate with AI scientists and leadership to optimize data quality, throughput, and cost, and define dataset roadmap for next-gen products.
Bachelor's, Master's, or PhD in Computer Science or related field.
5+ years of software development experience.
Proficiency with bash/Python scripting in Linux environments.
Experience with Docker, Infrastructure-as-Code, and at least one major cloud provider (GCP preferred).
Experienced in designing and operating large-scale data ingestion and processing workflows in cloud environments.
Comfortable working collaboratively with AI researchers to improve dataset quality and efficiency.
Able to manage multiple priorities in a fast-growing, distributed startup setting with minimal supervision.