← Zurück zur Liste
Stelle · Mid-level

AI Adversarial Specialist - Fully Remote | Up to $62/hr

Sonstige • Mid-level • Remote • Vollzeit Deutschland Deutschland

Mercor is hiring an English/Finnish bilingual AI Safety expert for contract, fully remote red-teaming work to uncover vulnerabilities in conversational AI models.

Responsibilities

  • Red team conversational AI models and agents to identify jailbreaks, prompt injections, misuse cases and bias exploitation
  • Generate high-quality human data by annotating failures, classifying vulnerabilities and flagging systemic risks
  • Apply structured approaches using taxonomies, benchmarks and playbooks for consistent testing
  • Document findings reproducibly to produce reports, datasets and attack cases
  • Work independently and asynchronously to meet deadlines while improving AI model performance

Requirements

  • Fluent in English and Finnish
  • Prior experience in red teaming, AI adversarial work, cybersecurity or socio-technical probing
  • Ability to communicate risks clearly to technical and non-technical stakeholders

Nice to have

  • Experience with adversarial ML, including jailbreak datasets, prompt injection and model extraction
  • Background in cybersecurity, such as penetration testing and exploit development
  • Expertise in socio-technical risk, including harassment/disinfo probing and abuse analysis

Soft skills

Independent and asynchronous workCommunicating risk to technical and non-technical audiences

What we offer

  • Contract pay of $48-62/hour
  • Fully remote

About the company

Mercor connects elite creative and technical talent with leading AI research labs; headquartered in San Francisco, backed by investors including Benchmark, General Catalyst, Peter Thiel, Adam D'Angelo, Larry Summers and Jack Dorsey.

Languages: Angol: Felsőfok, Finn: Felsőfok

Ähnliche Stellen