Source Job

Canada

  • Conduct red-team evaluations to identify jailbreaks, prompt injections, and misuse scenarios in conversational AI models.
  • Develop creative adversarial prompts and scenarios to systematically probe model behavior and uncover weaknesses.
  • Generate high-quality human evaluation data by annotating failures and classifying vulnerabilities.

Cybersecurity Penetration Testing Creative Writing

20 jobs similar to AI Safety Expert — English & Swedish

Jobs ranked by similarity.

Canada

  • Conduct adversarial testing on conversational AI to identify vulnerabilities.
  • Generate human evaluation data and document findings reproducibly.
  • Communicate risks to both technical and non-technical stakeholders.

AI safety evaluation company that tests and improves conversational AI systems for vulnerabilities and biases. Operates with a remote-first culture and collaborates on diverse projects to promote responsible AI development.

Canada

  • Conduct hands-on adversarial testing across AI models, applications, and data pipelines to identify vulnerabilities.
  • Perform advanced red-team assessments including jailbreaks, guardrail bypass, and prompt injection analysis.
  • Produce detailed vulnerability reports and collaborate with AI safety and engineering teams to improve security.

This company specializes in AI security research and adversarial machine learning. It is a remote-first organization with a collaborative culture focused on technical impact and professional growth.

Canada

  • Apply deep subject-matter expertise to AI model evaluation and large language model projects.
  • Develop challenging domain-specific problems and assess AI responses for accuracy and reasoning.
  • Collaborate with AI research teams to improve training datasets and evaluation methodologies.

Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. It offers a remote, asynchronous work culture and uses AI tools to support recruitment.

US

  • Deliver AI red teaming security assessments for enterprise clients, defining scopes and authoring detailed reports.
  • Build and improve frameworks and tooling to scale AI red teaming service delivery and automate repetitive tasks.
  • Develop novel red teaming methodologies for emerging AI modalities and stay ahead of the latest AI security threats.

Check Point protects over 100,000 organizations worldwide from cyber and AI-driven threats using a prevention-first approach. They are recognized by TIME, Newsweek, and Forbes for excellence and workplace culture.

US

  • Evaluate AI-generated content for quality, accuracy, and cultural relevance
  • Apply Castilian Spanish expertise to assess response appropriateness for Spain
  • Provide structured feedback and document decisions to improve AI performance

Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. They use objective, data-driven recruitment processes and prioritize privacy and fairness.

$15–$15/hr
United States

  • Rating and assessing the performance of AI models based on their output or behavior.
  • Labeling and categorizing content to train machine learning models.
  • Generating prompts, responses, and summaries to improve language model reasoning.

Innodata is a global data engineering company that provides data, evaluation frameworks, and human expertise for AI systems. With over 36 years of experience, the company focuses on enabling responsible AI advancement.

Global

  • Provide native-level Canadian French language vetting and QA for AI data projects.
  • Annotate and review AI outputs for grammatical accuracy, cultural context, and naturalness.
  • Develop educational resources and feedback documentation to improve AI alignment.

We are an AI training company that focuses on language alignment and data annotation for AI systems. Our remote team values linguistic precision and cultural nuance.

$22–$22/hr
Canada

  • Evaluate and rank model outputs, stress-test models for failure modes, and create high-quality datasets with detailed rubrics.
  • Annotate and correct multimodal data, maintain consistency through calibration exercises, and adapt to evolving task types.
  • Report on model performance trends and provide clear feedback to cross-functional partners on model successes and failures.

Cohere is a security-first enterprise AI company that builds cutting-edge foundation models and end-to-end products for real-world business problems. It is a global technology company with offices in Toronto, San Francisco, London, New York, Montreal, Seoul, Germany, and Paris, staffed by a team of passionate researchers, engineers, and designers.

Canada

  • Evaluate AI-generated documents, spreadsheets, and presentation decks against quality rubrics.
  • Identify factual, formatting, visual, and structural issues in professional deliverables.
  • Provide clear, structured feedback to enhance AI output quality and consistency.

This partner company specializes in AI training and evaluation, focusing on improving the quality of AI-generated professional content. Operating as a remote and asynchronous team, they value precision, collaboration, and independent work.

$4,000–$4,500/mo
Poland

  • Design and iterate complex system prompts and chain-of-thought structures for consumer AI experiences.
  • Partner with engineers and product managers to optimize prompt specifications for latency and cost.
  • Develop evaluation frameworks and playbooks to guardrail LLM outputs against bias and hallucination.

BOLD is a global company that creates digital products to help people build resumes, cover letters, and CVs, empowering job seekers in 180 countries. They are an established organization that values diversity and inclusion, with a culture of growth and professional fulfillment.

Brazil

  • Test and evaluate AI chatbots and language models through structured conversations using assigned criteria.
  • Assess AI-generated responses for quality, relevance, safety, and linguistic accuracy.
  • Submit accurate deliverables such as written evaluations, ratings, and audio recordings within required timelines.

This company specializes in AI development and data evaluation, focusing on improving generative AI systems. The organization operates with a flexible, project-based team and values linguistic expertise.

UK

  • Compare and rank AI-generated responses for accuracy, logic, and safety.
  • Review CS research papers alongside AI summaries to ensure scientific integrity.
  • Fact-check technical data and code for logical flaws and inaccuracies.

Prolific is building the largest pool of quality human data in the world, serving over 35,000 AI developers and researchers. They connect researchers with paid participants to gather high-quality, ethically sourced behavioral data for AI development.

US

  • Review, evaluate, and annotate AI-generated content across text, images, audio, and video.
  • Perform quality checks to ensure accuracy, consistency, and compliance with project guidelines.
  • Identify edge cases and inconsistencies, contribute to high-quality dataset development, and participate in calibration activities.

Welo Data, part of Welocalize, is a global AI data company with over 500,000 contributors that provides high-quality, ethical data for training advanced AI systems. The company supports a diverse, global community across 100+ countries and offers project-based freelance opportunities with flexibility and growth potential.

Global

  • Evaluate prompts and AI-generated outputs for accuracy, clarity, and cultural appropriateness.
  • Review and correct text, analyze multimedia content, and contribute voice recordings.
  • Apply local insight into tone, symbolism, visual cues, and market fit to deliver culturally relevant content.

LILT is an AI company that makes the world's information available to everyone, no matter the language they speak. They work with a global community of linguists and subject matter experts to deliver multilingual AI and human-verified services to Enterprises, Governments, and AI Developers.

Global

  • Evaluate LLM architecture logic for technical accuracy and audit ML code and notebooks for efficiency.
  • Refine RLHF frameworks to align models with human intent and analyze model reasoning in complex chain-of-thought prompts.
  • Benchmark performance by conducting comparative testing between model outputs based on technical metrics.

Prolific connects researchers with a global pool of participants for collecting high-quality human data to train AI models. With over 35,000 users, they focus on ethical data gathering to advance AI capabilities.

Canada

  • Evaluate AI-generated legal and business documents against quality standards and apply professional judgment.
  • Review contracts, diligence materials, redlines for accuracy, consistency, and completeness.
  • Provide clear, structured feedback to improve AI-generated legal content.

Global

  • Evaluate LLM responses for accuracy, clarity, and completeness.
  • Fact-check technical claims using authoritative references.
  • Validate code and outputs, and annotate model performance.

Prolific builds the largest pool of high-quality human data for AI development, serving over 35,000 AI developers, researchers, and organizations. They connect researchers with a global community to collect ethically sourced behavioral data.

UK

  • Evaluate AI-generated French responses, rate them, and flag cultural issues.
  • Rewrite weak responses into clear, natural Canadian French.
  • Create original French prompts and example responses to expand training data.

We are a global AI data company that delivers high-quality, ethical data to train the world's most advanced AI systems. With over 500,000 contributors, we offer flexible, remote project-based opportunities with a supportive global community.

Canada

  • Evaluate AI-generated documents and presentations against quality standards.
  • Apply humanities expertise to identify inaccuracies and cultural issues.
  • Provide structured feedback to improve AI model performance.

A partner company is seeking a humanities evaluator to assess AI-generated content for accuracy and quality. The company emphasizes cultural awareness and critical thinking in a remote, asynchronous work environment.

US

  • Utilize automatic prompt generation tools to create baseline prompts for complex parent-child template clusters.
  • Run and supervise automated prompt optimization, review outputs, and flag deadlocks or plateaus.
  • Manually draft, test, and refine prompts to handle edge cases, anti-patterns, and solve complex template architectures.

Welo Data provides AI services focused on data validation and model evaluation. They operate remotely with a global team and emphasize technical expertise and collaboration.