Source Job

Netherlands

  • Label and evaluate photos, graphics, videos, stickers, and designs in Dutch for linguistic accuracy and cultural appropriateness.
  • Assess AI-generated content against Canva's quality bar for Dutch users to shape localised AI experiences.
  • Build and contribute to Dutch-specific datasets and deliver labelled assets on time across varied task types.

Dutch AI Evaluation Quality Assurance Detail-oriented

20 jobs similar to AI Quality Evaluator - Dutch

Jobs ranked by similarity.

Egypt

  • Label, annotate, and evaluate Arabic photos, graphics, videos, stickers, and designs for linguistic quality and cultural appropriateness.
  • Assess AI-generated content against Canva’s quality bar for Arabic-speaking users.
  • Build and maintain Arabic-specific datasets to support AI feature internationalization.

Canva is an online graphic design platform that enables users to create visual content like social media graphics and presentations. With thousands of employees globally, the company fosters a culture of innovation, collaboration, and inclusion.

Global

  • Write prompts in Luxembourgish to test AI models and evaluate responses using a structured rating guide.
  • Identify areas for improvement in AI responses and provide accurate English translations.
  • Work remotely on a flexible schedule with competitive pay.

CrowdGen by Appen is a platform that connects freelancers to help improve AI models through data annotation and evaluation. It is part of a large global company with a diverse community of independent contractors working remotely.

Global

  • Train and evaluate cutting-edge AI models by completing language tasks in Italian/Spanish/French/German/Dutch.
  • Judge the performance of AI in performing Italian prompts and improve its capabilities.
  • Analyze, edit, and write in target languages with strong attention to detail for up to one hour per task.

Prolific is building the largest pool of quality human data in the world for AI training. Over 35,000 AI developers, researchers, and organizations use the platform to gather data from paid participants with diverse experiences, skills, and knowledge.

Global

  • Provide linguistic QA and develop alignment resources for AI data projects.
  • Review, annotate, and test AI outputs for grammatical accuracy and cultural context.
  • Act as a primary quality check to identify and correct subtle errors in Czech language.

We specialize in AI data projects, focusing on linguistic quality and cultural relevance. We are a global project that operates remotely with freelance contractors.

Sweden

  • Review completed tasks from trainers, including questions, images, and golden answers, to ensure accuracy and consistency.
  • Independently verify golden answers based on images and flag errors, inconsistencies, or ambiguities.
  • Provide clear feedback and track error patterns to maintain high data quality and escalate unclear cases.

Welo Data, part of Welocalize, is a global AI data company with over 500,000 contributors delivering high-quality, ethical data to train advanced AI systems. They operate in 100+ countries, offering flexible remote work and growth opportunities.

  • Evaluate prompts and AI-generated outputs for accuracy, cultural appropriateness, and brand alignment.
  • Review and correct text, analyze multimedia content, and contribute voice recordings.
  • Apply local cultural insight and consistent evaluation guidelines to ensure high-quality AI training.

Lilt provides multilingual AI and human-verified services to enterprises, governments, and AI developers. They foster a global community of linguists and subject matter experts working on cutting-edge AI and language technology.

  • Review online ads and provide human feedback to improve ad quality and relevance.
  • Evaluate ad matching, language accuracy, and cultural relevance for German-speaking markets.
  • Contribute your human perspective to AI training, working at your own pace.

TELUS Digital enhances AI via human intelligence by providing enriched data through a skilled team of AI specialists and a managed AI community of over one million crowd contributors. They maintain a diverse international culture and support remote work with flexible hours.

Turkey

  • Annotate and review multimedia data (video, images, metadata) using defined labeling rules and guidelines.
  • Perform self-QA, track recurring issues, and contribute to guideline improvements with clear documentation.
  • Collaborate with stakeholders to meet throughput and quality targets while participating in calibration sessions.

Welo Data provides AI services, specializing in data annotation and quality review for multimedia content. As a growing remote team, we offer freelance opportunities for detail-oriented professionals to enhance global AI systems.

Japan

  • Evaluate the Japanese app and web experience across core user journeys to identify language, terminology, and cultural issues.
  • Update translations directly in the translation management system and document issues that require broader product or design changes.
  • Collaborate with the Localization team to prioritize improvements and ensure consistency in Japanese terminology.

Whatnot is the largest live shopping platform in North America and Europe, enabling users to buy, sell, and discover items through live video. They are a remote co-located team with hubs across multiple countries, recently named the #1 Best Startup Employer in America by Forbes.

Global

  • Evaluate AI-generated text and audio in Catalan for accuracy and natural flow.
  • Provide corrections and constructive feedback on grammar, tone, and cultural context.
  • Complete approximately 10 hours of asynchronous tasks each week via our online platform.

Terac is building the world's largest pool of vetted human experts for AI. Researchers, AI labs, and product teams use Terac to recruit, screen, and pay study participants across industries, languages, and skill sets.

India

  • Evaluate AI-generated responses for accuracy, grammar, and cultural relevance.
  • Identify issues and provide refined, high-quality rewritten responses.
  • Create natural prompts and responses in Hindi to improve conversational datasets.

Welo Data, part of Welocalize, is a global AI data company with 500,000+ contributors delivering high-quality, ethical data to train the world’s most advanced AI systems. They're building smarter, more human AI with a diverse community in 100+ countries.

Global

  • Evaluate AI-generated text and voice snippets in Punjabi for naturalness and authenticity.
  • Assess audio clips for cultural and tonal accuracy of AI speech.
  • Provide feedback on linguistic nuance and quality of AI outputs.

Prolific builds the biggest pool of quality human data in the world, serving over 35,000 AI developers and researchers. The company connects researchers with paid study participants from diverse backgrounds to gather high-quality, ethically sourced behavioral data.

Global

  • Provide native-level Canadian French language vetting and QA for AI data projects.
  • Annotate and review AI outputs for grammatical accuracy, cultural context, and naturalness.
  • Develop educational resources and feedback documentation to improve AI alignment.

We are an AI training company that focuses on language alignment and data annotation for AI systems. Our remote team values linguistic precision and cultural nuance.

$24–$24/hr
Germany

  • Provide expert rating support and guidance, enhancing the skills of other raters on AI-powered advertising systems.
  • Analyze datasets, identify patterns, and use KPIs to drive data-informed decisions for quality improvements.
  • Collaborate with clients and teams to conduct test rounds, ensure quality benchmarks, and maintain rating proficiency.

Welo Data, part of Welocalize, is a global AI data company with 500,000+ contributors delivering high-quality, ethical data to train advanced AI systems. The company values limitless flexibility and growth, offering a supportive global community for its contributors.

Global

  • Evaluate side-by-side text and voice snippets to assess quality and authenticity of Marathi speech.
  • Listen to AI-generated audio and rate how natural the voice sounds.
  • Identify unnatural tone, pronunciation, or cultural mismatches in AI outputs.

Prolific is not just another player in the AI space – we are building the biggest pool of quality human data in the world. Over 35,000 AI developers, researchers, and organizations use Prolific to gather data from paid study participants with a wide variety of experiences, knowledge, and skills.

Global

  • Review English source documents alongside two machine-generated Assamese translations, evaluating accuracy, fluency, and overall quality.
  • Select the preferred translation and provide a clear written justification for your assessment.
  • Complete assigned samples independently within established timelines, adhering strictly to project and client guidelines.

Welo Data provides AI operations and data generation services for leading technology clients. They are a global contributor community offering project-based opportunities with flexible, remote work.

Global

  • Validate Dutch language codes and metadata for web pages, apps, and documents.
  • Perform screen reader testing using native-language tools to ensure translated content is accurate and natural.
  • Review multimedia elements like captions and audio descriptions for synchronization and proper labeling.

Welocalize, a Welo Global brand, serves localization teams through AI-enabled multilingual content solutions that enable enterprises to operate and scale globally. Welocalize combines AI, automation, and human expertise to support enterprises in more than 300 languages, enabling accurate, culturally aligned, and compliant multilingual content at scale.

Global

  • Evaluate AI-generated Telugu text for linguistic accuracy, grammar, and cultural relevance.
  • Rewrite and improve language outputs with natural prompts and responses.
  • Collaborate with global teams to enhance conversational AI model quality.

Welo Data, part of Welocalize, is a global AI data company with 500,000+ contributors delivering high-quality, ethical data to train advanced AI systems. They have a diverse community in 100+ countries and offer project-based opportunities with full remote autonomy.

Global

  • Write prompts in Slovak to test AI models.
  • Evaluate AI responses using a structured rating guide.
  • Translate prompts and evaluations into English.

CrowdGen by Appen provides AI training data services to improve machine learning models. They operate a global community of independent contractors and emphasize flexible, remote work.

$35–$45/hr
Global

  • Curate and annotate multilingual audio data to train AI for voice interactions and speech recognition.
  • Ensure high-quality voice recordings and accurate transcriptions across diverse languages and accents.
  • Collaborate with technical staff to improve annotation tools and audio workflows.

SpaceXAI creates AI systems to understand the universe and aid humanity in its pursuit of knowledge. The team is small, highly motivated, and focused on engineering excellence.