Mercor seeks an AI Safety Red Teamer to stress-test frontier AI models by designing adversarial prompts and identifying jailbreaks, unsafe behaviors, and policy failures. The role covers evaluating model robustness across misinformation, cyber, biosecurity, fraud, politics, and other sensitive domains.
Responsibilities include documenting vulnerabilities and contributing to safety benchmarking, while collaborating with AI researchers to improve alignment, robustness, and safety in a fully remote
#J-*****-Ljbffr
📌 Remote Ai Safety Red Team Expert For Frontier Models (Milano)
🏢 Altro
📍 Milano