AI Adversarial Specialist

M
MercorAI safety
Source API remote eligibility restrictions: IndiaContract
Salary20 - 22 USD per hour
Apply NowOpens the employer's application page

Job Details

Languages
Native fluency in English and Urdu.
Required Skills
Cybersecurity

Requirements

  • Demonstrate native fluency in English and Urdu.
  • Have prior experience in red teaming, AI adversarial work, cybersecurity, or socio-technical probing.
  • Be able to explain risks clearly to technical and non-technical stakeholders.
  • Adversarial ML, cybersecurity, or socio-technical risk analysis experience is preferred.
  • Creative-probing skills, such as psychology, acting, or writing, are preferred.

Responsibilities

  • Red-team conversational AI models and agents to identify jailbreaks, prompt injections, and misuse cases.
  • Annotate failures, classify vulnerabilities, and flag systemic risks.
  • Follow taxonomies, benchmarks, and playbooks to ensure consistent testing.
  • Document findings reproducibly in reports, datasets, and attack cases.
  • Review AI outputs on sensitive topics such as bias and misinformation.
View Full Description & ApplyYou'll be redirected to the employer's site
20 - 22 USD per hour
Apply Now