Source Job

$15–$15/hr
US

  • Review, evaluate, and validate AI-generated content and data according to project guidelines.
  • Compare AI outputs and identify the most accurate, relevant, or high-quality results.
  • Annotate, categorize, or label text, audio, images, video, or other data with clear feedback.

Generative AI Data Annotation English Critical Thinking

20 jobs similar to Generative AI Analyst

Jobs ranked by similarity.

US

  • Evaluate AI-generated content for quality, accuracy, and cultural relevance
  • Apply Castilian Spanish expertise to assess response appropriateness for Spain
  • Provide structured feedback and document decisions to improve AI performance

Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. They use objective, data-driven recruitment processes and prioritize privacy and fairness.

US

  • Review, evaluate, and annotate AI-generated content across text, images, audio, and video.
  • Perform quality checks to ensure accuracy, consistency, and compliance with project guidelines.
  • Identify edge cases and inconsistencies, contribute to high-quality dataset development, and participate in calibration activities.

Welo Data, part of Welocalize, is a global AI data company with over 500,000 contributors that provides high-quality, ethical data for training advanced AI systems. The company supports a diverse, global community across 100+ countries and offers project-based freelance opportunities with flexibility and growth potential.

$15–$15/hr
United States

  • Rating and assessing the performance of AI models based on their output or behavior.
  • Labeling and categorizing content to train machine learning models.
  • Generating prompts, responses, and summaries to improve language model reasoning.

Innodata is a global data engineering company that provides data, evaluation frameworks, and human expertise for AI systems. With over 36 years of experience, the company focuses on enabling responsible AI advancement.

Global

  • You will rate and assess the performance of AI models based on their output or behavior.
  • You will label elements of content and assign predefined categories to generate training data.
  • You will create prompts, summaries, and evaluate relevance to improve AI system understanding.

Innodata (Nasdaq: INOD) is a global data engineering company that provides data, evaluation frameworks, and human expertise for AI systems. The company has a 36+ year legacy of delivering high-quality data and outstanding outcomes for customers.

US

  • Engage in conversations with a real-time speech-to-speech AI model
  • Evaluate performance on speech recognition, audio quality, conversation flow, and content accuracy
  • Provide accurate ratings based on project guidelines within specified timelines

Appen is a global leader in AI training data and crowd-sourced solutions. They work with a large community of independent contractors to improve AI systems through human evaluation.

US

  • Perform side-by-side comparisons of AI-generated responses.
  • Evaluate outputs for accuracy, relevance, clarity, instruction-following, and overall quality.
  • Assess AI responses across general-purpose questions and answers, web search results, file-based and image-based responses, content-generation tasks, and single-turn and multi-turn conversations.

We are a technology solutions firm headquartered in Bellevue, Washington, with teams across the United States. We help organizations turn complex challenges into meaningful outcomes by connecting strategy and execution across AI, cloud, data, product development, and emerging technology.

Global

  • Evaluate prompts and AI-generated outputs for accuracy, clarity, and cultural appropriateness.
  • Review and correct text, analyze multimedia content, and contribute voice recordings.
  • Apply local insight into tone, symbolism, visual cues, and market fit to deliver culturally relevant content.

LILT is an AI company that makes the world's information available to everyone, no matter the language they speak. They work with a global community of linguists and subject matter experts to deliver multilingual AI and human-verified services to Enterprises, Governments, and AI Developers.

$4–$5/hr
Global

  • Review AI-generated responses against source images and quality guidelines.
  • Identify issues like hallucinations, missing details, or policy violations.
  • Provide structured feedback to improve model performance and output quality.

Jobgether uses AI-powered matching to connect candidates with partner companies. They focus on efficient, objective hiring processes and operate as a platform for remote opportunities.

Canada

  • Evaluate AI-generated documents, spreadsheets, and presentation decks against quality rubrics.
  • Identify factual, formatting, visual, and structural issues in professional deliverables.
  • Provide clear, structured feedback to enhance AI output quality and consistency.

This partner company specializes in AI training and evaluation, focusing on improving the quality of AI-generated professional content. Operating as a remote and asynchronous team, they value precision, collaboration, and independent work.

Global

  • Evaluate LLM responses for accuracy, clarity, and completeness.
  • Fact-check technical claims using authoritative references.
  • Validate code and outputs, and annotate model performance.

Prolific builds the largest pool of high-quality human data for AI development, serving over 35,000 AI developers, researchers, and organizations. They connect researchers with a global community to collect ethically sourced behavioral data.

Netherlands

  • Label and evaluate photos, graphics, videos, stickers, and designs in Dutch for linguistic accuracy and cultural appropriateness.
  • Assess AI-generated content against Canva's quality bar for Dutch users to shape localised AI experiences.
  • Build and contribute to Dutch-specific datasets and deliver labelled assets on time across varied task types.

Canva is redefining how the world experiences design, empowering users to create visual content. The company has a global team and supports flexible, remote-friendly work, with a focus on collaboration and innovation.

Canada

  • Evaluate AI-generated documents and presentations against quality standards.
  • Apply humanities expertise to identify inaccuracies and cultural issues.
  • Provide structured feedback to improve AI model performance.

A partner company is seeking a humanities evaluator to assess AI-generated content for accuracy and quality. The company emphasizes cultural awareness and critical thinking in a remote, asynchronous work environment.

Global

  • Write, rewrite, and rate generated content and conversations for AI training.
  • Translate and review content between Slovenian and English with cultural accuracy.
  • Annotate data and images related to current events, pop culture, and more.

They help businesses build, train, and manage AI systems. They offer flexible remote work with variable workload and weekly pay.

$22–$22/hr
Canada

  • Evaluate and rank model outputs, stress-test models for failure modes, and create high-quality datasets with detailed rubrics.
  • Annotate and correct multimodal data, maintain consistency through calibration exercises, and adapt to evolving task types.
  • Report on model performance trends and provide clear feedback to cross-functional partners on model successes and failures.

Cohere is a security-first enterprise AI company that builds cutting-edge foundation models and end-to-end products for real-world business problems. It is a global technology company with offices in Toronto, San Francisco, London, New York, Montreal, Seoul, Germany, and Paris, staffed by a team of passionate researchers, engineers, and designers.

Spain

  • Evaluate machine translations from English to Spanish (Spain) to improve translation engine accuracy.
  • Work as an independent contractor with flexible hours, completing files at your own pace.
  • Must be a native Spanish speaker with strong English proficiency and reside in Spain.

CrowdGen by Appen connects independent contractors with AI training projects, including machine translation evaluation. The company is part of Appen, a global leader in AI training data, and fosters a flexible, remote work culture.

Global

  • Evaluate AI-generated text and voice snippets in Marathi for quality and authenticity.
  • Listen to audio clips and rate how natural the AI voice sounds.
  • Provide feedback on tone, pronunciation, and cultural context.

Prolific is an AI data platform that connects researchers with a global pool of participants to gather high-quality, ethically sourced human data. With over 35,000 AI developers and organizations using the platform, Prolific is building the largest pool of quality human data to train AI models.

Global

  • Evaluate model-generated content across multiple modalities including text, images, audio, and video.
  • Apply defined quality rubrics such as factuality, consistency, and aesthetics to assess outputs.
  • Conduct independent research on unfamiliar topics to make well-supported evaluation judgments.

Our client is a global technology company that helps businesses build, train, and manage AI systems. They offer flexible, remote work with variable workload and a focus on high-quality model evaluation.

Global

  • Contribute to building smarter, more accurate AI by annotating, evaluating, and creating prompts for language models.
  • Work flexibly on your own terms with remote projects that fit your schedule and skills.
  • Be part of a global community of linguists and tech enthusiasts shaping the future of AI.

We are a global AI data company with 500,000+ contributors delivering high-quality, ethical data to train the world's most advanced AI systems. We are building a diverse community in 100+ countries, offering flexible remote work and limitless opportunities for growth.

United States

  • Review search results and evaluate their relevance to user queries
  • Answer true/false questions about content quality
  • Rate search results based on guidelines to improve AI systems

Welo Data provides AI services and data validation to improve search engine and AI systems. They are a remote-first company with a focus on quality and support for their contractors.

Sweden

  • Review completed tasks from trainers, including questions, images, and golden answers, to ensure accuracy and consistency.
  • Independently verify golden answers based on images and flag errors, inconsistencies, or ambiguities.
  • Provide clear feedback and track error patterns to maintain high data quality and escalate unclear cases.

Welo Data, part of Welocalize, is a global AI data company with over 500,000 contributors delivering high-quality, ethical data to train advanced AI systems. They operate in 100+ countries, offering flexible remote work and growth opportunities.