Source Job

Global

  • Evaluate AI-generated responses for safety and bias against strict rubrics.
  • Classify harmful content categories like hate speech and self-harm.
  • Verify factual accuracy of model claims using external sources to prevent hallucinations.

Content Moderation Data Annotation Fact-Checking Risk Assessment Policy Analysis

20 jobs similar to AI Safety & Policy Expert

Jobs ranked by similarity.

Global

  • Review text or media samples based on provided project guidelines
  • Apply accurate labels and categorizations to diverse data sets
  • Evaluate AI-generated responses for clarity, safety, and factual accuracy

Terac is building the world's largest pool of vetted human experts for AI. Researchers, AI labs, and product teams use Terac to recruit, screen, and pay study participants across industries, languages, and skill sets.

Australia

  • Evaluate AI-generated responses for relevance, accuracy, and personalization using personalized prompts and data from connected Google applications.
  • Identify subtle issues such as incorrect assumptions, irrelevant recommendations, inconsistencies, and inappropriate personalization.
  • Provide clear, detailed, and structured feedback to support improvements to AI models and personalization systems.

Our partner company is seeking an AI Response Quality Evaluator to improve AI-generated responses. This is a project-based contract role with a remote, independent working environment and a duration of up to 16 weeks.

Global 4w PTO

  • Review and approve or reject user-generated content based on platform standards, documenting moderation decisions.
  • Conduct safety and compliance reviews on in-house content and workflows, including datasets and AI characters.
  • Moderate video, image, and text content, taking action on flagged violations and categorizing content for appropriate visibility.

EverAI is building the world's largest AI companionship platform, driven by a proprietary moderation system called EverGuard. With 50 million users in two years and a fully remote team of about 100, the company is a fast-growing, category-creating AI company.

Global

  • Evaluate prompts and AI-generated outputs for accuracy, clarity, and cultural appropriateness.
  • Review and correct text, analyze multimedia content, and contribute voice recordings.
  • Apply careful judgment to ensure high-quality results aligned with task objectives.

LILT provides multilingual AI and human-verified services to enterprises, governments, and AI developers. The company has a global community of linguists and language professionals committed to innovation and excellence.

India

  • Annotate text, image, audio, and video data to train large language models and AI systems.
  • Evaluate search results, ads, and chatbot outputs for factual accuracy, linguistic quality, and safety.
  • Provide detailed feedback and documentation to refine AI model reasoning and evaluation frameworks.

Our partner is a company focused on AI training and data annotation, working with a global network of linguists and tech enthusiasts. It values flexibility, transparency, and cultural accuracy in AI development.

US

  • Craft creative prompts and multi-turn scenarios to stress-test AI guardrails across diverse risk categories.
  • Discover ways around safety filters and restrictions using jailbreak, evasion, and prompt injection techniques.
  • Evaluate and score model responses against structured harm taxonomies and severity rubrics.

Handshake AI partners with leading AI research labs to make models safer and more robust. Our red teaming operations help identify vulnerabilities before they reach users, contributing directly to the responsible development of frontier AI systems.

Global

  • Fact-check health information generated by AI to ensure accuracy and up-to-date content.
  • Review medical terminology and ensure proper use of names for bones, muscles, medicines, or conditions.
  • Flag any dangerous or incorrect health advice to maintain user safety.

TELUS Digital provides AI training data and services to help build better AI models. They manage a global community of over one million contributors from diverse backgrounds, offering flexible independent contractor opportunities.

Global

  • Review AI-generated business emails, reports, and strategies to ensure they are professionally sound.
  • Fact-check AI suggestions for marketing plans, budgets, and adherence to standard business rules.
  • Teach AI common office tasks such as meeting summaries, invoice drafting, and project timeline creation.

TELUS Digital operates a global AI community of over one million contributors who help collect, enhance, and train content for AI models. The community is diverse, flexible, and focused on shaping innovative AI technologies used by world-class brands.

Ireland

  • Annotate and label text, images, audio, or other content to support AI model training.
  • Evaluate AI-generated outputs such as search results and chatbot responses for quality and relevance.
  • Create and refine prompts while reviewing Greek linguistic and cultural accuracy.

This opportunity is a talent network for Greek-speaking contributors supporting AI model training through annotation, evaluation, and prompt creation. It is a global, remote community of independent contributors with flexible project-based work.

Germany

  • Label, annotate, and evaluate German-language content including photos, graphics, and videos for linguistic and cultural accuracy.
  • Evaluate AI-generated content against Canva's quality bar for German users to shape language experiences.
  • Build and contribute to German-specific datasets to support the internationalization of Canva AI features.

Canva is a design platform redefining how the world experiences design. It is a global company with a large user base, known for its innovative culture and focus on AI-powered features.

India

  • Analyze, evaluate, and review diverse datasets to support AI system training and improvement.
  • Assess AI-generated content for accuracy, relevance, consistency, and quality, providing actionable feedback.
  • Work independently in a remote, digital-first environment, managing multiple tasks and deadlines.

  • Assess the clarity, coherence, and accuracy of written content to ensure it meets project standards.
  • Conduct detailed writing evaluations and provide constructive feedback for continual improvement.
  • Identify and annotate AI-generated content, focusing on detecting low-quality or artificial text elements.

Our client is a rapidly growing, venture-backed AI company building intelligent systems with human expertise and machine learning workflows. Backed by more than $40 million in funding, the company connects a global network of experts to high-impact AI projects.

Global

  • Analyze multi-turn conversation flows to reduce cognitive load and ensure concise answers.
  • Refine AI's calibration of trust between overconfidence and caution.
  • Develop specialized personas for tone-alignment across industries.

TELUS Digital builds intuitive AI by refining human-AI interaction. They have a global community of over one million contributors and emphasize flexible, remote collaboration.

Global

  • Review AI-generated financial statements for mathematical accuracy and compliance with local laws.
  • Verify correct application of tax rates and rules for specific countries.
  • Help AI recognize suspicious patterns for fraud detection training.

TELUS Digital trains AI to handle financial data by reviewing AI-generated statements and tax advice. They have a global community of over one million contributors and offer flexible remote work on an independent contractor basis.

Germany

  • Annotate and label text, images, audio, or other data to improve AI systems.
  • Evaluate search results, advertisements, and chatbot responses for quality and accuracy.
  • Create, test, and refine prompts for large language models with Maltese-language insight.

Our partner is a global contributor network that provides flexible remote AI projects focused on annotation, evaluation, and prompt creation. The work is independent, project-based, and aims to improve the accuracy, relevance, and inclusivity of AI systems.

Global

  • Audit AI 'Chain-of-Thought' reasoning for medical diagnostics to ensure adherence to international clinical protocols.
  • Rigorously verify AI-generated drug interactions, dosage recommendations, and contraindications against peer-reviewed medical literature.
  • Precisely label complex biomedical datasets—including MRI/CT scans, pathology reports, and genomic data—to improve model visual-spatial reasoning.

TELUS Digital is a technology company that helps businesses collect, enhance, train, translate, and localize content to build better AI models. They have a global AI community of over one million contributors from diverse backgrounds, fostering a culture of flexibility and innovation.

Global

  • Review customer signals, user exchanges, and code to identify potential policy violations.
  • Enforce Safeguards workflows and policy guidelines with precision and care.
  • Collaborate with Strategy Analysts, Policy Managers, and Threat Intelligence Investigators to identify trends and improve safety.

We provide talent acquisition services, including building recruiting processes, attracting top talent, and payrolling contractors. Part of a family of brands alongside Flawless Recruit and Recruiter.com, we have a collaborative, fast-paced culture.

Spain

  • Evaluate search results and AI-generated content for quality, relevance, accuracy, and usefulness.
  • Conduct online research to verify information and support rating decisions.
  • Participate in training, calibration sessions, and ongoing quality reviews.

TELUS Digital is a global AI community that helps customers collect, enhance, train, translate, and localize content to build better AI models. It consists of a vibrant network of over 1 million contributors from diverse backgrounds.

US

  • Test and evaluate AI model responses across diverse topics and conversation types.
  • Write and refine system prompts to shape model behavior and personality.
  • Apply detailed rubrics consistently to assess model performance and document findings.

We are a team dedicated to evaluating and improving advanced AI models and products. Our fast-moving environment emphasizes quality, independent judgment, and attention to detail.

$15–$15/hr
Global

  • Evaluate and label AI model outputs to improve performance and alignment with project guidelines.
  • Create prompts, rewrite text, and generate training data for large language models.
  • Work on flexible, remote, project-based tasks while helping shape the future of AI.

Innodata is a global data engineering company that enables the responsible advancement of AI by providing data, evaluation frameworks, and human expertise. With a 36+ year legacy, the company delivers high-quality data and outstanding outcomes for customers.