





Remote, mid-level AI role with generalist senior title and metro Bangalore exposure increases applicant competition.
Core LLM and ML engineering skills are transferable, though public-health and Indic language experience raises domain specificity.
Mandatory 3+ years AI product experience, 1+ year LLM work, and specific evaluation/LLM skillset enforce strict filters.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Own the performance monitoring, evaluation, and improvement of voice AI agents for public health applications serving millions of low-income users.
Develop and maintain automated evaluation pipelines, simulation environments, and annotation workflows for multilingual datasets.
Collaborate closely with NGO and government partners including field visits to translate requirements into production-ready AI solutions.
3+ years experience building AI products used by real users, with at least 1 year on LLM-powered products.
Strong foundation in math, deep learning, LLMs, and modern AI systems.
Proficient in Python programming and hands-on development of AI evaluation systems.
Location: Remote, with occasional travel to Bangalore.
Operates with a product mindset prioritizing real-world usability and equitable safe performance over pure research artifacts.
Experienced in rigorous quantitative evaluation methods and systematic failure mode analysis.
Able to translate complex research into scalable, production-level AI pipelines and iterate rapidly based on field data and user feedback.