AI Safety Expert - Adversarial ML
New
M
MercorAI Research
United StatesContract
Salary16 - 22 USD per hour
Apply NowOpens the employer's application page
Job Details
- Languages
- English, Odia
- Required Skills
- CybersecurityPrompt Engineering
Requirements
- Native fluency in English and Odia.
- Strong judgment about language and content.
- Rigorous attention to detail and consistency.
- Structured approach to guidelines and quality standards.
- Clear communication with technical and non-technical audiences.
- Adaptability across projects, task types, and customers.
Responsibilities
- Red team conversational AI models and agents.
- Conduct jailbreaks, prompt injections, misuse cases, and bias exploitation.
- Generate high-quality human data.
- Annotate failures, classify vulnerabilities, and flag systemic risks.
- Follow taxonomies, benchmarks, and playbooks to ensure consistent testing.
- Produce reports, datasets, and attack cases that customers can act on.
- Work independently and asynchronously.
View Full Description & ApplyYou'll be redirected to the employer's site