AI Adversarial Specialist
M
MercorAI safety
Source API remote eligibility restrictions: IndiaContract
Salary20 - 22 USD per hour
Apply NowOpens the employer's application page
Job Details
- Languages
- Native fluency in English and Urdu.
- Required Skills
- Cybersecurity
Requirements
- Demonstrate native fluency in English and Urdu.
- Have prior experience in red teaming, AI adversarial work, cybersecurity, or socio-technical probing.
- Be able to explain risks clearly to technical and non-technical stakeholders.
- Adversarial ML, cybersecurity, or socio-technical risk analysis experience is preferred.
- Creative-probing skills, such as psychology, acting, or writing, are preferred.
Responsibilities
- Red-team conversational AI models and agents to identify jailbreaks, prompt injections, and misuse cases.
- Annotate failures, classify vulnerabilities, and flag systemic risks.
- Follow taxonomies, benchmarks, and playbooks to ensure consistent testing.
- Document findings reproducibly in reports, datasets, and attack cases.
- Review AI outputs on sensitive topics such as bias and misinformation.
View Full Description & ApplyYou'll be redirected to the employer's site