
AI Safety & Red Team Expert | $29–45/hr | Worldwide Remote
Mercor is assembling a team of adversarial AI experts to probe, stress-test, and improve the safety of cutting-edge AI models. This role is ideal for professionals with backgrounds in AI red teaming, cybersecurity, or socio-technical risk analysis. Native fluency in both English and Portuguese (global, excluding Brazilian Portuguese) is required.
The safest AI is one that has already been challenged. Mercor's red team generates the adversarial human data that helps customers build more robust, trustworthy AI systems. You'll surface vulnerabilities that automated tests miss and deliver reproducible findings that drive real safety improvements.
Note: This project involves reviewing AI outputs on sensitive topics such as bias, misinformation, and harmful behaviors. All work is text-based. Participation in higher-sensitivity projects is optional and supported by clear guidelines and wellness resources. Topics are communicated in advance.