Mercor is assembling a panel of radiological safety experts to red-team frontier AI models. The goal is to test whether a model can correctly judge the misuse potential of a technical request—answering legitimate questions fully while refusing genuinely dangerous ones.
You will write prompts at three levels (benign, dual-use, adversarial), evaluate model responses, and craft the reference answer with detailed reasoning for how to respond correctly.
#J-18808-Ljbffr
📌 Radiation Safety Officer — AI Model Red-Teaming Lead (Toronto)
🏢 Obsidian
📍 Toronto
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.