Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey .
Position: AI Safety Red Teamer Type: Contract Location: Remote Design adversarial prompts to stress-test frontier AI models . Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains. Document vulnerabilities and contribute to safety benchmarking and red-teaming reports.
Collaborate with AI researchers to improve model alignment, robustness, and safety. Bachelor's degree or higher in Computer Science , Cybersecurity , Journalism , Communications , Psychology , Biology , Chemistry , Public Policy , or a related discipline. ~5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety,
cybersecurity, investigative journalism, life sciences, or a related field. ~ Experience designing adversarial prompts or evaluating frontier AI systems .
Experience with AI Red Teaming , RLHF , SFT , AI Alignment , or Trust & Safety . Familiarity with jailbreak testing, prompt engineering, or adversarial evaluation methodologies. Expertise in one or more grey-area domains, including cyber, biosecurity, political content, misinformation, or scientific safety.
AI interview based on your resume For details about the interview process and platform information, please check: Please complete your AI interview and application steps to be considered for this opportunity. #
📌 Ai safety specialist - remote (Roma)
🏢 Mercor
📍 Roma