AI Evaluation Specialist at Mercor (Toronto)

AI Evaluation Specialist at Mercor (Toronto)

26 Aug
|
Obsidian
|
Toronto

26 Aug

Obsidian

Toronto

Leverage your PhD expertise as an AI Evaluation Specialist at Mercor. Collaborate with top AI labs to create groundbreaking scientific evaluation tasks that challenge current models.
Mercor seeks scientists with a PhD or Master's to craft AI evaluation benchmarks in scientific computing. You will author original research problems in domains like numerical linear algebra and computational finance. This role combines mathematical depth with practical coding skills, requiring a commitment of 20+ hours per week over six weeks.
Key Responsibilities:
• Source material from papers, datasets, or custom scenarios
• Write scientific prompts based on established inputs
• Develop grading criteria for correct answers
• Calibrate tasks against leading AI models
• Ensure tasks are challenging enough for frontier models
Requirements:
• PhD in mathematics or closely related field
• Depth in two subdomains like computational finance
• Proficiency in Python or R for scientific tasks
• Familiar with Git and Docker workflows
• Bonus: Publications in peer-reviewed journals
Utilize your analytical and coding skills to advance AI research at Mercor.
#J-18808-Ljbffr

📌 AI Evaluation Specialist at Mercor (Toronto)
🏢 Obsidian
📍 Toronto

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: ai evaluation specialist at mercor (toronto) / toronto

Subscribe to this job alert:

Get the latest job offers by email for: ai evaluation specialist at mercor (toronto) / toronto