Source Job

Canada

  • Conduct adversarial testing on conversational AI to identify vulnerabilities.
  • Generate human evaluation data and document findings reproducibly.
  • Communicate risks to both technical and non-technical stakeholders.

Cybersecurity Penetration Testing Norwegian

9 jobs similar to AI Safety Expert — English & Norwegian

Jobs ranked by similarity.

Canada

  • Conduct red-team evaluations to identify jailbreaks, prompt injections, and misuse scenarios in conversational AI models.
  • Develop creative adversarial prompts and scenarios to systematically probe model behavior and uncover weaknesses.
  • Generate high-quality human evaluation data by annotating failures and classifying vulnerabilities.

The partner company is a technology organization focused on AI safety and responsible AI development. They work with a remote, asynchronous team to improve the robustness of conversational AI systems.

Canada

  • Conduct hands-on adversarial testing across AI models, applications, and data pipelines to identify vulnerabilities.
  • Perform advanced red-team assessments including jailbreaks, guardrail bypass, and prompt injection analysis.
  • Produce detailed vulnerability reports and collaborate with AI safety and engineering teams to improve security.

This company specializes in AI security research and adversarial machine learning. It is a remote-first organization with a collaborative culture focused on technical impact and professional growth.

UK

  • Evaluate AI-generated French responses, rate them, and flag cultural issues.
  • Rewrite weak responses into clear, natural Canadian French.
  • Create original French prompts and example responses to expand training data.

We are a global AI data company that delivers high-quality, ethical data to train the world's most advanced AI systems. With over 500,000 contributors, we offer flexible, remote project-based opportunities with a supportive global community.

US

  • Deliver AI red teaming security assessments for enterprise clients, defining scopes and authoring detailed reports.
  • Build and improve frameworks and tooling to scale AI red teaming service delivery and automate repetitive tasks.
  • Develop novel red teaming methodologies for emerging AI modalities and stay ahead of the latest AI security threats.

Check Point protects over 100,000 organizations worldwide from cyber and AI-driven threats using a prevention-first approach. They are recognized by TIME, Newsweek, and Forbes for excellence and workplace culture.

Global

  • Evaluate prompts and AI-generated outputs for accuracy, clarity, and cultural appropriateness.
  • Review and correct text, analyze multimedia content, and contribute voice recordings.
  • Apply local insight into tone, symbolism, visual cues, and market fit to deliver culturally relevant content.

LILT is an AI company that makes the world's information available to everyone, no matter the language they speak. They work with a global community of linguists and subject matter experts to deliver multilingual AI and human-verified services to Enterprises, Governments, and AI Developers.

US

  • Evaluate AI-generated content for quality, accuracy, and cultural relevance
  • Apply Castilian Spanish expertise to assess response appropriateness for Spain
  • Provide structured feedback and document decisions to improve AI performance

Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. They use objective, data-driven recruitment processes and prioritize privacy and fairness.

$15–$15/hr
United States

  • Rating and assessing the performance of AI models based on their output or behavior.
  • Labeling and categorizing content to train machine learning models.
  • Generating prompts, responses, and summaries to improve language model reasoning.

Innodata is a global data engineering company that provides data, evaluation frameworks, and human expertise for AI systems. With over 36 years of experience, the company focuses on enabling responsible AI advancement.

Canada

  • Apply deep subject-matter expertise to AI model evaluation and large language model projects.
  • Develop challenging domain-specific problems and assess AI responses for accuracy and reasoning.
  • Collaborate with AI research teams to improve training datasets and evaluation methodologies.

Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. It offers a remote, asynchronous work culture and uses AI tools to support recruitment.

Canada

  • Evaluate AI-generated documents and presentations against quality standards.
  • Apply humanities expertise to identify inaccuracies and cultural issues.
  • Provide structured feedback to improve AI model performance.

A partner company is seeking a humanities evaluator to assess AI-generated content for accuracy and quality. The company emphasizes cultural awareness and critical thinking in a remote, asynchronous work environment.