Contractual Role - 3 months (will be extended depending on performance and project requirement)
Location - Hyderabad (Nanakramguda)
Key Responsibilities
Test and validate AI-generated insights, recommendations, and decision-making workflows.
Evaluate LLM and RAG systems for accuracy, relevance, consistency, factuality, and hallucinations.
Validate retrieval quality, context relevance, grounding, and response quality in RAG systems.
Test AI agents and autonomous workflows across functional, negative, and edge-case scenarios.
Define AI evaluation criteria, test datasets, quality metrics, and validation processes.
Perform regression testing for models, prompts, RAG configurations, and AI workflows.
Collaborate with AI/ML engineers to identify issues and improve AI system quality
Required Skills
Solid understanding of AI/ML and Generative AI testing.
2 to 3+ years hands-on experience testing LLM and RAG-based applications.
Hands-on LLM testing (GPT, Claude, Gemini, Llama)
Experience testing at least one RAG-based application
Hallucination, relevance, groundedness and response quality validation
Solid analytical and problem-solving skills.
📌 Ai Quality Engineer Contractual Hyderabad
🏢 SIDE
📍 Hyderabad
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.