Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey .
Position: AI Safety Experts — English & Danish
Type: Contract
Compensation: $48–$62/hour
Location: Remote
Role Responsibilities
- Red team conversational AI models and agents to identify jailbreaks, prompt injections, misuse cases, and bias exploitation.
- Generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks.
- Apply structure by following taxonomies, benchmarks, and playbooks to maintain consistent testing.
- Document reproducibly by producing reports, datasets, and attack cases that customers can act on.
- Work independently and asynchronously to meet deadlines while improving AI model performance.
Qualifications
Must-Have
- Fluent in English and Danish.
- Prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing.
- Solid communication skills to explain risks clearly to both technical and non-technical stakeholders.
Preferred
- Experience in Adversarial ML, Cybersecurity, or socio-technical risk analysis.
- Skills in creative probing such as psychology, acting, or writing for unconventional adversarial thinking.