





Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Entry-level GenAI contractor in Bengaluru at a VC-backed startup draws high applicant density.
ML/LLM evaluation skills moderately transferable, favoring candidates with specific LLM and data experience.
Moderate filters: degree enrollment plus Python/NLP familiarity but no strict years or certifications.
Build and operate evaluation systems for LLM-powered chatbot performance, implementing Evaluation-Driven Development processes across teams.
Design and calibrate LLM-as-judge scoring methods with subject matter experts to align automated and human assessments.
Support real-time monitoring tooling and analyze production chatbot interactions to identify and drive improvements.
Currently pursuing a Bachelor's or Master's degree (B.E., B.Tech, M.Tech, or equivalent) in Computer Science, Software Engineering, or related field from an Indian university.
Familiarity with Python and experience working with APIs, data pipelines, or scripting.
Based in Bengaluru, India with availability for hybrid work (2 days in office per week).
Work Experience Required: Not explicitly mentioned in the JD.
Demonstrates strong analytical thinking and attention to detail, focused on verifying model performance before deployment.
Has curiosity and foundational knowledge about large language models, AI evaluation, and related NLP concepts possibly through coursework or personal projects.
Comfortable collaborating with ML engineers, analysts, and SMEs to iterate on prompts, models, and evaluation frameworks.