AI Safety Expert

New
J
JobgetherAI Safety
CanadaContract
Salary48 - 62 USD per hour
Apply NowOpens the employer's application page

Job Details

Languages
Fluent or professional-level proficiency in both English and Swedish
Required Skills
Cybersecurity

Requirements

  • Fluent or professional-level proficiency in both English and Swedish.
  • Prior experience in AI red teaming, adversarial AI research, cybersecurity, penetration testing, or socio-technical probing.
  • Demonstrated ability to systematically test complex systems for vulnerabilities, unexpected behaviors, and potential misuse.
  • Strong understanding of structured testing approaches, including the use of taxonomies, benchmarks, frameworks, or established playbooks.
  • Excellent analytical and critical-thinking skills, combined with creativity and curiosity when exploring unconventional attack paths.
  • Strong ability to document technical findings clearly and explain risks to both technical and non-technical audiences.
  • High attention to detail and commitment to producing reproducible, evidence-based evaluations.
  • Comfortable working independently in a remote, asynchronous environment and adapting quickly across different projects.

Responsibilities

  • Conduct red-team evaluations of conversational AI models and agents to identify jailbreaks, prompt-injection vulnerabilities, misuse cases, and potential bias exploitation.
  • Develop creative adversarial prompts and scenarios to systematically probe model behavior and uncover weaknesses.
  • Generate high-quality human evaluation data by annotating model failures, classifying vulnerabilities, and flagging systemic safety risks.
  • Apply established taxonomies, benchmarks, frameworks, and testing playbooks to ensure consistent and rigorous evaluations.
  • Document findings in a reproducible manner through detailed reports, datasets, attack cases, and supporting evidence.
  • Communicate identified risks and technical findings clearly to both technical and non-technical stakeholders.
  • Adapt testing strategies across different AI systems, projects, use cases, and requirements while maintaining consistent evaluation standards.
  • Work independently and asynchronously while meeting deadlines and maintaining high-quality evaluation standards.
View Full Description & ApplyYou'll be redirected to the employer's site
48 - 62 USD per hour
Apply Now