AI Safety Expert
New
J
JobgetherAI Safety
CanadaContract
Salary48 - 62 USD per hour
Apply NowOpens the employer's application page
Job Details
- Languages
- Fluent or professional-level proficiency in both English and Swedish
- Required Skills
- Cybersecurity
Requirements
- Fluent or professional-level proficiency in both English and Swedish.
- Prior experience in AI red teaming, adversarial AI research, cybersecurity, penetration testing, or socio-technical probing.
- Demonstrated ability to systematically test complex systems for vulnerabilities, unexpected behaviors, and potential misuse.
- Strong understanding of structured testing approaches, including the use of taxonomies, benchmarks, frameworks, or established playbooks.
- Excellent analytical and critical-thinking skills, combined with creativity and curiosity when exploring unconventional attack paths.
- Strong ability to document technical findings clearly and explain risks to both technical and non-technical audiences.
- High attention to detail and commitment to producing reproducible, evidence-based evaluations.
- Comfortable working independently in a remote, asynchronous environment and adapting quickly across different projects.
Responsibilities
- Conduct red-team evaluations of conversational AI models and agents to identify jailbreaks, prompt-injection vulnerabilities, misuse cases, and potential bias exploitation.
- Develop creative adversarial prompts and scenarios to systematically probe model behavior and uncover weaknesses.
- Generate high-quality human evaluation data by annotating model failures, classifying vulnerabilities, and flagging systemic safety risks.
- Apply established taxonomies, benchmarks, frameworks, and testing playbooks to ensure consistent and rigorous evaluations.
- Document findings in a reproducible manner through detailed reports, datasets, attack cases, and supporting evidence.
- Communicate identified risks and technical findings clearly to both technical and non-technical stakeholders.
- Adapt testing strategies across different AI systems, projects, use cases, and requirements while maintaining consistent evaluation standards.
- Work independently and asynchronously while meeting deadlines and maintaining high-quality evaluation standards.
View Full Description & ApplyYou'll be redirected to the employer's site