Mercor seeks experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing. You will design challenging prompts, uncover model weaknesses, and evaluate AI behavior across complex, high-risk topics.Responsibilities include designing prompts to stress-test models, identifying jailbreaks and policy failures, and documenting vulnerabilities for safety benchmarking. Collaboration with researchers will improve alignment, robustness, and safety.
#J-18808-Ljbffr
📌 Frontier AI Safety Red Team Specialist (Remote) (Madrid)
🏢 Mercor
📍 Madrid