AI Safety Expert - Adversarial ML

New
M
MercorAI Research
United StatesContract
Salary16 - 22 USD per hour
Apply NowOpens the employer's application page

Job Details

Languages
English, Odia
Required Skills
CybersecurityPrompt Engineering

Requirements

  • Native fluency in English and Odia.
  • Strong judgment about language and content.
  • Rigorous attention to detail and consistency.
  • Structured approach to guidelines and quality standards.
  • Clear communication with technical and non-technical audiences.
  • Adaptability across projects, task types, and customers.

Responsibilities

  • Red team conversational AI models and agents.
  • Conduct jailbreaks, prompt injections, misuse cases, and bias exploitation.
  • Generate high-quality human data.
  • Annotate failures, classify vulnerabilities, and flag systemic risks.
  • Follow taxonomies, benchmarks, and playbooks to ensure consistent testing.
  • Produce reports, datasets, and attack cases that customers can act on.
  • Work independently and asynchronously.
View Full Description & ApplyYou'll be redirected to the employer's site
16 - 22 USD per hour
Apply Now