Back to Jobs
Jobgether
AI & Machine Learning 3d ago

AI Safety Expert — English & Norwegian

Jobgether
CanadaCanada
Contract
$48–$62 per hour
Mid-Level

Job Description

Key Skills Required

Master these to land this role

Machine LearningBestseller 🔥
Learn in 42 Hours
Prompt EngineeringBestseller 🔥
Learn in 22 Hours
CybersecurityAI EngineerNLP

Want to know if you're a match for this job?

Calculate My Match Score

This contract role offers an opportunity to help strengthen the safety and reliability of conversational AI systems and intelligent agents. You will conduct structured adversarial testing to uncover vulnerabilities, jailbreaks, prompt injections, misuse scenarios, and potential bias exploitation. Your work will help identify weaknesses before they can affect users or real-world deployments. You will generate high-quality human evaluation data by annotating failures, classifying vulnerabilities, and identifying systemic risks. The role combines creative problem-solving with rigorous testing methodologies, frameworks, and benchmarks. Working remotely across evolving projects, you will produce reproducible findings that help technical and non-technical stakeholders take meaningful action.

Accountabilities:

  • Conduct red-team testing of conversational AI models and agents to identify jailbreaks, prompt-injection vulnerabilities, misuse scenarios, and opportunities for bias exploitation.
  • Develop creative and realistic adversarial prompts and attack scenarios designed to probe model weaknesses.
  • Generate high-quality human evaluation data by annotating model failures, classifying vulnerabilities, and identifying systemic safety risks.
  • Apply established taxonomies, benchmarks, testing frameworks, and playbooks to ensure consistent and rigorous evaluations.
  • Document findings in a reproducible manner through detailed reports, datasets, attack cases, and supporting evidence.
  • Clearly communicate technical findings and risk assessments to both technical and non-technical stakeholders.
  • Adapt testing approaches across different projects, AI systems, use cases, and customer requirements while maintaining consistent quality standards.
  • Contribute insights that help improve AI safety practices, model robustness, and risk mitigation strategies.

Requirements:

  • Fluent or professional-level proficiency in both English and Norwegian, with strong written and verbal communication skills.
  • Prior experience in AI red teaming, adversarial AI work, cybersecurity, penetration testing, or socio-technical risk assessment.
  • Demonstrated ability to systematically probe complex systems for vulnerabilities, unexpected behaviors, and potential misuse.
  • Experience applying structured frameworks, taxonomies, benchmarks, or testing methodologies to security or safety evaluations.
  • Strong analytical and critical-thinking skills, combined with the creativity needed to develop novel probing strategies.
  • Ability to clearly document technical findings and communicate them effectively to both technical and non-technical audiences.
  • Strong attention to detail and commitment to producing reproducible, high-quality evaluation results.
  • Adaptability and ability to move efficiently between different projects, models, use cases, and customer requirements.
  • Experience in adversarial machine learning, cybersecurity, or socio-technical risk is preferred.
  • Creative-probing skills developed through psychology, acting, creative writing, or related disciplines are an advantage.

Benefits:

  • Competitive contract compensation of $48–$62 per hour.
  • Fully remote work environment.
  • Flexible, project-based working structure.
  • Opportunity to contribute directly to the safety and responsible development of advanced AI systems.
  • Exposure to diverse AI models, agents, safety challenges, and adversarial testing methodologies.
  • Opportunity to apply both technical and creative problem-solving skills to emerging AI safety challenges.
  • Work across varied projects and use cases with the opportunity to expand expertise in AI safety and security.
  • Meaningful contribution to improving the robustness, reliability, and responsible deployment of conversational AI.

How would you rate this job post?

See what other professionals think about this role.

banner

Jobgether (operating under jobgether.com) is a prominent, AI-powered global remote job discovery engine and workforce solution platform engineered to eliminate borders in the recruitment landscape. Founded in 2020 by a team of serial recruitment and tech entrepreneurs, the platform leverages advanced data collection, machine learning, and automated web parsing technologies to aggregate, curate, and verify hundreds of thousands of active flexible, hybrid, and fully remote employment listings from thousands of companies worldwide. Jobgether addresses the major challenges of geographical sourcing mismatch and hidden location constraints (such as timezone requirements or country-specific tax residency limitations) by indexing jobs based on detailed "remoteness profiles." The platform offers a clean, candidate-first experience featuring tailored matching agents, real-time alerts, and granular search parameters, helping professionals easily discover authentic borderless career paths. For enterprises, Jobgether functions as a high-efficiency global branding and talent acquisition channel, expanding their employer brand footprint, driving high-intent candidate matching loops, and tapping into a highly diverse international talent pool without the standard friction of multi-jurisdictional recruitment pipelines. Headquartered in Brussels, Belgium, Jobgether operates as a leading technological utility connecting remote-first software engineers, digital marketers, product leaders, and operations specialists directly with innovative employers across the globe.

Safety First

  • Never pay for a job application.
  • Do not share sensitive bank info.
  • Verify the client before starting work.
Learn More