Mercor
$29–45/hr
163 people hired
You will red team conversational AI models and agents to uncover vulnerabilities, misuse cases, bias, and unsafe behaviors. | The role involves adversarial testing through jailbreaks, prompt injections, multi-turn manipulation, and socio-technical probing. | You will generate structured human data and reproducible attack cases that help strengthen AI safety systems. | Native fluency in English and Portuguese (excluding Brazilian Portuguese) is required.
Bilingual AI Safety Expert