





Niche AI-testing role but metro Hyderabad increases applicant supply modestly.
LLM/RAG testing expertise is highly domain-specific, limiting cross-industry transferability.
Requires mandatory hands-on LLM/RAG testing and AI-evaluation skills, significantly narrowing eligible candidates.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Test and validate AI-generated insights, recommendations, and workflows focusing on accuracy, relevance, hallucinations, and consistency.
Evaluate LLM and RAG systems including retrieval quality, context relevance, and grounding.
Define evaluation criteria and perform regression testing; collaborate with AI/ML engineers to improve AI system quality.
Hands-on experience testing LLM and RAG-based applications.
Strong understanding of AI/ML and Generative AI testing.
Work Mode: Onsite in Hyderabad, 5 days a week (WFO).
Work Experience Required: Not explicitly mentioned in the JD.
Experienced in identifying and mitigating hallucinations and quality regressions in AI outputs.
Comfortable working with AI evaluation metrics and developing automated testing frameworks.
Capable of collaborating closely with AI/ML engineers to enhance AI system reliability and trustworthiness.