AI Safety Expert - Red Teaming
Mercor · United States · 2 days ago
RemoteRemoteResearch$48–$62/hrPart-time
Role Responsibilities
- Red team conversational AI models and agents to identify jailbreaks, prompt injections, misuse cases, and bias exploitation.
- Generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks.
- Apply structure by following taxonomies, benchmarks, and playbooks to maintain consistent testing.
- Document reproducibly by producing reports, datasets, and attack cases that customers can act on.
- Work independently and asynchronously to meet deadlines while improving AI model performance.
Qualifications
- Must-have: Fluent in English and Danish.
- Prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing.
- Strong communication skills to explain risks clearly to both technical and non-technical stakeholders.
Preferred Experience
- In Adversarial ML, Cybersecurity, or socio-technical risk analysis.
- Skills in creative probing such as psychology, acting, or writing for unconventional adversarial thinking.