About The Job Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey .
Position: AI Safety Experts — English & Tamil
Type: Contract Compensation: $16–$22/hour Location: Remote Role Responsibilities
- Red team conversational AI models and agents through jailbreaks, prompt injections, and misuse cases. Identify bias exploitation and multi-turn manipulation.
- Generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks.
- Apply structure by following taxonomies, benchmarks, and playbooks to ensure consistent testing.
- Document reproducibly by producing reports, datasets, and attack cases that customers can act on.
- Work independently and asynchronously to meet deadlines while improving AI model performance.
Qualifications Must-Have
- Fluent Language Skills Required: English & Tamil. Native fluency in English and Tamil is required.
- Strong judgment about language and content accuracy, completeness, and appropriateness.
- Rigorous attention to subtle errors, inconsistencies, and gaps.
- Structured approach to work, adhering to guidelines and quality standards.
- Explicit communication with both technical and non-technical audiences.
- Adaptability across projects, task types, and customers.
Preferred
- Experience in Adversarial ML: jailbreak datasets, prompt injection, RLHF/DPO attacks, model extraction.
- Cybersecurity skills: penetration testing, exploit development, reverse engineering.
- Knowledge of socio-technical risk: harassment/disinfo probing, abuse analysis, conversational AI testing.
- Creative probing skills: psychology, acting, writing for unconventional adversarial thinking.
Application Process (Takes 20–30 mins to complete)
- Upload resume
- AI interview based on your resume
- Submit form
Resources & Support
- For details about the interview process and platform information, please check: https://talent.docs.mercor.com/welcome
- For any help or support, reach out to:
[email protected]
PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity. ,
📌 AI Safety Expert - Adversarial ML (India)
🏢 Mercor
📍 India