Technical AI Evaluation Specialist (3 Months Contract) (India)

Technical AI Evaluation Specialist (3 Months Contract) (India)

08 Sep
|
Innodata
|
India

08 Sep

Innodata

India

Technical AI Evaluation Specialist (Code & Software Engineering)

EXPERIENCE: 2+ Years Software Dev or Solid CS Degree

LOCATION: Remote

LANGUAGE BAR: C1 Equivalent (Technical Writing & Code Review)

EMPLOYMENT: Contract 3 Months

1. ROLE OVERVIEW & OBJECTIVE

We are seeking rigorous Technical AI Evaluation Specialists to benchmark, evaluate, and align frontier Large Language Models (LLMs) specialized in code generation, multi-turn technical reasoning, and software architecture. In this role, you will analyze model-generated code against strict correctness, complexity, security, and UI fidelity standards across text-to-code, image-to-code, and side-by-side (SxS) comparison tasks. You will be responsible for uncovering subtle failure modes, edge-case vulnerabilities, and producing evidence-based, defensible rationales to guide model fine-tuning and Reinforcement Learning from Human Feedback (RLHF).

2. KEY RESPONSIBILITIES & CORE WORKFLOWS

- Model-Generated Code Evaluation: Evaluate AI-generated code for syntactical validity, execution accuracy, algorithmic complexity, and architectural best practices across diverse languages.
- Image-to-Code & UI Verification: Assess model capability in rendering pixel-perfect, responsive front-end components from wireframes, mockups, and UI design screenshots.
- Text-to-Code & Pairwise Analysis: Perform rigorous side-by-side (SxS) evaluations to determine model preference, scoring completions against granular multi-dimensional rubrics.
- Failure Mode & Edge Case Reasoning: Stress-test model completions against corner cases, boundary conditions, race conditions, memory leaks,



and input sanitization vulnerabilities.
- Evidence-Based Rationales: Write authoritative, structured C1-level technical rationales explaining exact point deductions, execution trace errors, and counterfactual fixes.
- Guideline Calibration & Feedback: Collaborate with research engineers and prompt authors to refine evaluation rubrics, establish baseline test harnesses, and identify emerging model degradation patterns.

3. CANDIDATE PROFILE & QUALIFICATIONS

✓ Mandatory Requirements

• Experience: 2+ years of professional software development experience OR a strong Computer Science / Software Engineering degree with demonstrable coding proficiency.

- Front-End / UI Exposure: Hands-on experience translating design mockups to responsive front-end code (HTML5, modern CSS, React, or TypeScript frameworks).
- Code Review & RLHF Exposure: Proven background in structured code reviews, automated testing, or prior experience in AI model evaluation/RLHF pipelines.

★ Preferred Qualifications

- Multi-Language Fluency: Strong capability in Python, JavaScript/TypeScript, Java, C++, or Go, with an aptitude for rapidly reading and debugging unfamiliar frameworks.
- DevOps & Tooling: Familiarity with Git, Docker, CI/CD pipelines, AST parsers, or static code analysis tools.
- Advanced CS Foundations: Deep understanding of data structures, algorithms, concurrency, and secure coding practices (OWASP Top 10).
- Language & Communication: C1-equivalent professional English proficiency with exceptional technical articulation and defensive writing discipline.

📌 Technical AI Evaluation Specialist (3 Months Contract) (India)
🏢 Innodata
📍 India

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: technical ai evaluation specialist (3 months contract) (india) / india

Subscribe to this job alert:

Get the latest job offers by email for: technical ai evaluation specialist (3 months contract) (india) / india