Jobs · Research

AI Safety Expert - Red Teaming

Mercor · United States · 2 days ago
RemoteRemoteResearch$48–$62/hrPart-time

Role Responsibilities

  • Red team conversational AI models and agents to identify jailbreaks, prompt injections, misuse cases, and bias exploitation.
  • Generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks.
  • Apply structure by following taxonomies, benchmarks, and playbooks to maintain consistent testing.
  • Document reproducibly by producing reports, datasets, and attack cases that customers can act on.
  • Work independently and asynchronously to meet deadlines while improving AI model performance.

Qualifications

  • Must-have: Fluent in English and Danish.
  • Prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing.
  • Strong communication skills to explain risks clearly to both technical and non-technical stakeholders.

Preferred Experience

  • In Adversarial ML, Cybersecurity, or socio-technical risk analysis.
  • Skills in creative probing such as psychology, acting, or writing for unconventional adversarial thinking.

Similar jobs

AI Red Team Engineer

White CircleUnited States· 3 wk ago
RemoteEngineering$60k–$90k/yrapply on jobs.ashbyhq.com