← Back to list
Job · Mid-level

AI Adversarial Specialist - Fully Remote | Upto $62/hr

Other • Mid-level • Remote • Full-time Sweden Sweden

Mercor is hiring an AI Adversarial Specialist (AI Safety Expert) for a fully remote contract role, red-teaming conversational AI models to uncover jailbreaks, prompt injections, misuse cases, and bias exploitation. The role pays $48-62 per hour and requires fluency in English and Swedish.

Responsibilities

  • Red team conversational AI models and agents to identify jailbreaks, prompt injections, misuse cases, and bias exploitation
  • Generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks
  • Apply structure by following taxonomies, benchmarks, and playbooks to ensure consistent testing
  • Document findings reproducibly by producing reports, datasets, and attack cases
  • Work independently and asynchronously to meet deadlines while improving AI model performance

Requirements

  • Fluent in English and Swedish
  • Prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing
  • Strong communication skills to explain risks to technical and non-technical stakeholders

Nice to have

  • Experience with Adversarial ML, cybersecurity, and socio-technical risk
  • Skills in creative probing such as psychology, acting, or unconventional adversarial thinking

Soft skills

Strong communication skills for explaining risk to technical and non-technical audiencesAbility to work independently and asynchronouslyCreative, unconventional adversarial thinking

What we offer

  • Fully remote contract role
  • Compensation of $48-62 per hour

About the company

Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, its investors include Benchmark, General Catalyst, Peter Thiel, Adam D'Angelo, Larry Summers, and Jack Dorsey.

Similar jobs