We are looking for exceptional PhDs, researchers, scientists, and domain experts to contribute to an advanced AI Science Benchmarking Project.
This project focuses on designing and evaluating high-complexity scientific and technical problems to test the reasoning and problem-solving capabilities of modern AI systems.
• Design highly challenging, expert-level scientific or technical problems
• Create problems requiring several hours or more for a qualified expert to solve
• Develop structured and defensible solution approaches
• Define assumptions, constraints, edge cases, and expected outcomes
• Evaluate whether leading AI models can successfully solve these problems
• Identify where and why AI reasoning fails
• Review problems for correctness, ambiguity, consistency, and difficulty
• Refine problems to ensure clearly defensible or uniquely correct answers
Who We're Looking For:
• PhD preferred in a relevant scientific, mathematical, technical, or quantitative discipline
• Solid Master's candidates with demonstrated expertise may also be considered
• Strong research, academic, or professional subject-matter expertise
• Excellent analytical and structured problem-solving skills
• Ability to communicate complex concepts clearly
• Basic familiarity with AI tools and Large Language Models (LLMs)
There is no fixed minimum number of years of experience. Selection will primarily depend on your depth of expertise, quality of reasoning, and ability to create genuinely challenging problems.
Interested candidates can apply or message me directly on my email at [HIDDEN TEXT] with their updated CV, area of specialization, and current availability.
📌 AI Science Benchmarking Expert (India)
🏢 MyRemoteTeam
📍 India
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.