Remote Adversarial ML Specialist — AI Safety Red Team
Santa Clarita, Los Angeles County, California, 91350, USA
Listed on 2026-09-25
-
IT/Tech
AI Evaluation, Data Annotation/ AI Labeling
Mercor is building a remote red team for AI safety. You will lead adversarial testing of conversational AI models, focusing on jailbreaks, prompt injections, misuse, and bias exploitation.
You will generate high-quality human data, annotate failures, classify vulnerabilities, and document findings in reproducible reports and datasets that help customers strengthen their AI systems. Ideal candidates bring prior red teaming experience, a curious adversarial mindset, structured approaches, and the
Step into the Remote Adversarial ML Specialist — AI Safety Red Team role at Obsidian in San Francisco, CA, United States and grow with us.
This is a genuine role to take on the Remote Adversarial ML Specialist — AI Safety Red Team role at Obsidian.
As a Remote Adversarial ML Specialist — AI Safety Red Team, you will play an important part at Obsidian in San Francisco, CA, United States.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).