Back to Jobs
Jobgether
AI & Machine Learning 3d ago

AI Safety Expert — English & Swedish

Jobgether
CanadaCanada
Contract
$48–$62 per hour
Mid-Level

Job Description

Key Skills Required

Master these to land this role

Machine LearningBestseller 🔥
Learn in 42 Hours
Prompt EngineeringBestseller 🔥
Learn in 22 Hours
CybersecurityAI EngineerNLP

Want to know if you're a match for this job?

Calculate My Match Score

This contract role offers an opportunity to help identify and address safety risks in conversational AI models and intelligent agents. You will conduct structured red-team testing to uncover jailbreaks, prompt injections, misuse scenarios, and potential bias exploitation. Your work will help surface vulnerabilities and systemic risks before they impact users or real-world deployments. You will generate high-quality human evaluation data by annotating failures and classifying different types of vulnerabilities. The role combines rigorous testing methodologies with creative adversarial thinking and strong analytical judgment. Working remotely and asynchronously, you will produce actionable findings that contribute directly to safer, more robust, and more reliable AI systems.

Accountabilities:

  • Conduct red-team evaluations of conversational AI models and agents to identify jailbreaks, prompt-injection vulnerabilities, misuse cases, and potential bias exploitation.
  • Develop creative adversarial prompts and scenarios to systematically probe model behavior and uncover weaknesses.
  • Generate high-quality human evaluation data by annotating model failures, classifying vulnerabilities, and flagging systemic safety risks.
  • Apply established taxonomies, benchmarks, frameworks, and testing playbooks to ensure consistent and rigorous evaluations.
  • Document findings in a reproducible manner through detailed reports, datasets, attack cases, and supporting evidence.
  • Communicate identified risks and technical findings clearly to both technical and non-technical stakeholders.
  • Adapt testing strategies across different AI systems, projects, use cases, and requirements while maintaining consistent evaluation standards.
  • Work independently and asynchronously while meeting deadlines and maintaining high-quality evaluation standards.

Requirements:

  • Fluent or professional-level proficiency in both English and Swedish, with strong written and verbal communication skills.
  • Prior experience in AI red teaming, adversarial AI research, cybersecurity, penetration testing, or socio-technical probing.
  • Demonstrated ability to systematically test complex systems for vulnerabilities, unexpected behaviors, and potential misuse.
  • Strong understanding of structured testing approaches, including the use of taxonomies, benchmarks, frameworks, or established playbooks.
  • Excellent analytical and critical-thinking skills, combined with creativity and curiosity when exploring unconventional attack paths.
  • Strong ability to document technical findings clearly and explain risks to both technical and non-technical audiences.
  • High attention to detail and commitment to producing reproducible, evidence-based evaluations.
  • Comfortable working independently in a remote, asynchronous environment and adapting quickly across different projects.
  • Experience with adversarial machine learning, cybersecurity, or socio-technical risk is preferred.
  • Creative probing skills, including experience with psychology, acting, creative writing, or unconventional adversarial thinking, are an advantage.

Benefits:

  • Competitive contract compensation of $48–$62 per hour.
  • Fully remote work environment.
  • Flexible, asynchronous working structure.
  • Opportunity to contribute directly to the safety and responsible development of advanced AI systems.
  • Exposure to diverse conversational AI models, agents, safety challenges, and adversarial testing methodologies.
  • Opportunity to combine technical expertise with creative and unconventional problem-solving approaches.
  • Work across varied AI safety projects and use cases while expanding expertise in adversarial testing.
  • Meaningful contribution to improving AI model robustness, reliability, and responsible deployment.

How would you rate this job post?

See what other professionals think about this role.

banner

Jobgether (operating under jobgether.com) is a prominent, AI-powered global remote job discovery engine and workforce solution platform engineered to eliminate borders in the recruitment landscape. Founded in 2020 by a team of serial recruitment and tech entrepreneurs, the platform leverages advanced data collection, machine learning, and automated web parsing technologies to aggregate, curate, and verify hundreds of thousands of active flexible, hybrid, and fully remote employment listings from thousands of companies worldwide. Jobgether addresses the major challenges of geographical sourcing mismatch and hidden location constraints (such as timezone requirements or country-specific tax residency limitations) by indexing jobs based on detailed "remoteness profiles." The platform offers a clean, candidate-first experience featuring tailored matching agents, real-time alerts, and granular search parameters, helping professionals easily discover authentic borderless career paths. For enterprises, Jobgether functions as a high-efficiency global branding and talent acquisition channel, expanding their employer brand footprint, driving high-intent candidate matching loops, and tapping into a highly diverse international talent pool without the standard friction of multi-jurisdictional recruitment pipelines. Headquartered in Brussels, Belgium, Jobgether operates as a leading technological utility connecting remote-first software engineers, digital marketers, product leaders, and operations specialists directly with innovative employers across the globe.

Safety First

  • Never pay for a job application.
  • Do not share sensitive bank info.
  • Verify the client before starting work.
Learn More