← Zurück zur Liste
Stelle · Senior

AI Safety Specialist – Remote

Sonstige • Senior • Remote • Vollzeit • Estland Estland

Contract, remote role red-teaming frontier AI models with adversarial prompts and documenting the vulnerabilities uncovered.

Responsibilities

  • ▹Design adversarial prompts to stress-test frontier AI models
  • ▹Identify jailbreaks, unsafe behaviors, hallucinations and policy failures
  • ▹Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content and other sensitive domains
  • ▹Document vulnerabilities and contribute to safety benchmarking and red-teaming reports
  • ▹Collaborate with AI researchers to improve model alignment, robustness and safety

Requirements

  • ▹Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy or a related field
  • ▹5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism or life sciences
  • ▹Strong analytical reasoning, prompt design skills and written communication
  • ▹Experience designing adversarial prompts or evaluating frontier AI systems

Nice to have

  • ▹Experience with AI Red Teaming, RLHF, SFT, AI Alignment or Trust & Safety
  • ▹Familiarity with jailbreak testing, prompt engineering or adversarial evaluation methodologies
  • ▹Expertise in grey-area domains such as cyber, biosecurity, political content, misinformation or scientific safety

About the company

Mercor connects elite creative and technical talent with leading AI research labs; headquartered in San Francisco, its investors include Benchmark, General Catalyst, Peter Thiel, Adam D'Angelo, Larry Summers and Jack Dorsey.

Education: BSc vagy magasabb végzettség informatika, kiberbiztonság, újságírás, kommunikáció, pszichológia, biológia, kémia vagy közpolitika területén

Ähnliche Stellen