1. Role OverviewMercor is partnering with a leading AI research organization to engage experienced data scientists for a project focused on evaluating how well AI systems perform real-world data science work. Rather than producing deliverables yourself, you will define what excellent work looks like: designing task-specific grading criteria and scoring completed work samples with rigorous, well-reasoned written justifications.2. Key ResponsibilitiesDesign precise, task-specific grading criteria for real-world data science deliverables (analyses, models, dashboards, experiment readouts, and written recommendations)Score AI-generated and human work samples against those criteria, with detailed written justifications for every scoreApply consistent, evidence-based judgment so that scores are reproducible and defensibleIncorporate structured feedback from senior reviewers and iterate quickly on your work3.
Ideal Qualifications5+ years of skilled data science experience in industryBackground in business operations, product, or growth data science at top-tier technology companiesDeep fluency in experiment design and A/B testing, metric definition, SQL/Python analysis, and communicating findings to executive stakeholdersExceptionally strong written communicationDetail-oriented, consistent, and comfortable having your judgment reviewed and calibrated against peersPrior experience with AI training, evaluation, or human-data projects is a strong plus4. Application ProcessQualified applicants may be asked to complete a brief technical assessment or submit additional information #J-18808-Ljbffr
📌 Data Science Expert - Ai Evaluation (Toronto)
🏢 Mercor
📍 Toronto
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.