Mercor is hiring PhD and Master's scientists to author AI evaluation tasks (Sci Code) for a current benchmark in scientific computing. You will craft original, executable research problems that frontier models currently struggle to solve.
Domains require depth in at least two subfields with a coding focus; Python or R for scientific computing is essential. The role involves sourcing material, writing prompts, and building robust grading criteria for model evaluation.
#J-18808-Ljbffr
📌 AI Benchmark Architect for Scientific Computing (Quebec City)
🏢 Mercor
📍 Quebec City
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.