AI Safety Experts — English & Swedish
JobgetherRemotely
machine learningcybersecurityaiframeworkstaxonomies
Job Description
📋 Description
- Conduct red-team evaluations of conversational AI models and agents to identify jailbreaks
- Develop creative adversarial prompts and scenarios to systematically probe model behavior and
- Generate high-quality human evaluation data by annotating model failures, classifying
- Apply established taxonomies, benchmarks, frameworks, and testing playbooks to ensure consistent
- Document findings in a reproducible manner through detailed reports, datasets, attack cases, and
- Communicate identified risks and technical findings clearly to both technical and non-technical
🎯 Requirements
- Fluent or professional-level proficiency in both English and Swedish, with strong written and
- Prior experience in AI red teaming, adversarial AI research, cybersecurity, penetration testing, or
- Demonstrated ability to systematically test complex systems for vulnerabilities, unexpected
- Strong understanding of structured testing approaches, including the use of taxonomies, benchmarks
- Excellent analytical and critical-thinking skills, combined with creativity and curiosity when
- Strong ability to document technical findings clearly and explain risks to both technical and
🎁 Benefits
- Competitive contract compensation of $48–$62 per hour.
- Fully remote work environment.
- Flexible, asynchronous working structure.
- Opportunity to contribute directly to the safety and responsible development of advanced AI systems.
- Exposure to diverse conversational AI models, agents, safety challenges, and adversarial testing
- Opportunity to combine technical expertise with creative and unconventional problem-solving
Back to all jobs