





Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Tier-1 employer and Pune metro increase competition, but specialized Staff ML focus limits applicant pool.
Role requires deep ML infrastructure, LLM, and platform experience, so background fit is highly domain-specific.
Explicit 8+ years requirement plus mandatory LLM, Kubernetes, cloud, and ML infrastructure skills make shortlisting highly strict.
Lead architecture and delivery of Zendesk's GenAI platform including LLM Proxy, benchmarking, cost-control, and orchestration across product lines.
Own design and scaling of evaluation frameworks to gate model releases and define company-wide standards for safety and reasoning evaluation.
Drive platform reliability, observability, capacity planning, and lead cross-team initiatives on ML safety, risk, and quality trade-offs, mentoring senior engineers.
8+ years industry experience in backend, platform, or ML infrastructure engineering with major production responsibilities.
BS in Computer Science, Engineering, or related field, or equivalent practical experience.
Proficiency with Python (or similar), Kubernetes, cloud infrastructure (AWS/GCP/Azure), and ML/LLM production systems.
Location requirement: Must be physically located and able to work onsite part-time in Pune, India (Karnataka or Maharashtra).
Experienced in building large-scale distributed ML systems and infrastructure with a deep understanding of LLMs and inference serving.
Demonstrated ability to lead cross-functional technical projects and set company-wide ML safety and quality standards.
Strong system design skills focusing on scalable, reliable, cost-optimized platform solutions with stakeholder collaboration experience.