Source Job

US

  • Perform side-by-side comparisons of AI-generated responses.
  • Evaluate outputs for accuracy, relevance, clarity, instruction-following, and overall quality.
  • Assess AI responses across general-purpose questions and answers, web search results, file-based and image-based responses, content-generation tasks, and single-turn and multi-turn conversations.

English Annotation Analytical Thinking Attention To Detail

20 jobs similar to Castilian Spanish Labeler / Annotator

Jobs ranked by similarity.

Spain

  • Evaluate machine translations from English to Spanish (Spain) to improve translation engine accuracy.
  • Work as an independent contractor with flexible hours, completing files at your own pace.
  • Must be a native Spanish speaker with strong English proficiency and reside in Spain.

CrowdGen by Appen connects independent contractors with AI training projects, including machine translation evaluation. The company is part of Appen, a global leader in AI training data, and fosters a flexible, remote work culture.

US

  • Engage in conversations with a real-time speech-to-speech AI model
  • Evaluate performance on speech recognition, audio quality, conversation flow, and content accuracy
  • Provide accurate ratings based on project guidelines within specified timelines

Appen is a global leader in AI training data and crowd-sourced solutions. They work with a large community of independent contractors to improve AI systems through human evaluation.

Global

  • Evaluate AI-generated text and audio in Catalan for accuracy and natural flow.
  • Provide corrections and constructive feedback on grammar, tone, and cultural context.
  • Complete approximately 10 hours of asynchronous tasks each week via our online platform.

Terac is building the world's largest pool of vetted human experts for AI. Researchers, AI labs, and product teams use Terac to recruit, screen, and pay study participants across industries, languages, and skill sets.

Spain

  • Lead and support a team of approximately 10 Quality Control Reviewers, ensuring consistent quality standards across the Spanish (Spain) locale.
  • Act as the primary quality point of contact for the Spanish (Spain) locale, resolving quality questions and escalating complex issues when necessary.
  • Conduct calibration sessions and provide structured coaching to improve reviewer performance.

Welo Data provides data quality services for AI training programs, focusing on linguistic quality and annotation. The team is dedicated to improving AI systems through rigorous quality assurance and collaboration.

Global

  • Write prompts in Luxembourgish to test AI models and evaluate responses using a structured rating guide.
  • Identify areas for improvement in AI responses and provide accurate English translations.
  • Work remotely on a flexible schedule with competitive pay.

CrowdGen by Appen is a platform that connects freelancers to help improve AI models through data annotation and evaluation. It is part of a large global company with a diverse community of independent contractors working remotely.

Global

  • Write prompts in Slovak to test AI models.
  • Evaluate AI responses using a structured rating guide.
  • Translate prompts and evaluations into English.

CrowdGen by Appen provides AI training data services to improve machine learning models. They operate a global community of independent contractors and emphasize flexible, remote work.

Global

  • Side-by-side evaluation of text and voice snippets to assess quality and authenticity.
  • Naturalness assessment of AI-generated audio to rate how natural the voice sounds.
  • Quality control to identify where the AI's tone or pronunciation feels unnatural or culturally mismatched.

Prolific builds the largest pool of quality human data for AI development. Over 35,000 AI developers, researchers, and organizations use Prolific to gather data from paid study participants.

Global

  • Evaluate prompts and AI-generated outputs for accuracy, clarity, and cultural appropriateness.
  • Review and correct text, analyze multimedia content, and contribute voice recordings.
  • Apply local insight into tone, symbolism, visual cues, and market fit to deliver culturally relevant content.

LILT is an AI company that makes the world's information available to everyone, no matter the language they speak. They work with a global community of linguists and subject matter experts to deliver multilingual AI and human-verified services to Enterprises, Governments, and AI Developers.

Global

  • Provide native-level Canadian French language vetting and QA for AI data projects.
  • Annotate and review AI outputs for grammatical accuracy, cultural context, and naturalness.
  • Develop educational resources and feedback documentation to improve AI alignment.

We are an AI training company that focuses on language alignment and data annotation for AI systems. Our remote team values linguistic precision and cultural nuance.

$35–$45/hr
Global

  • Curate and annotate multilingual audio data to train AI for voice interactions and speech recognition.
  • Ensure high-quality voice recordings and accurate transcriptions across diverse languages and accents.
  • Collaborate with technical staff to improve annotation tools and audio workflows.

SpaceXAI creates AI systems to understand the universe and aid humanity in its pursuit of knowledge. The team is small, highly motivated, and focused on engineering excellence.

Argentina

  • Review and grade internet advertisements to help shape ad delivery based on user keywords.
  • Work remotely with a flexible schedule, minimum 5 hours per week, up to 20 hours.
  • Use both English and Spanish to evaluate ads and provide feedback.

Welo Data is an award-winning localization and data transformation company that runs one of the world’s largest Ads Rating Programs. The company offers a multicultural and international environment with opportunities for professional development and human support 24/6.

Ireland

  • Listen to two audio recordings and compare them to determine which is better.
  • Follow provided evaluation guidelines to make consistent judgments.
  • Complete approximately 20-23 cases per hour with flexible remote work.

Appen leverages human feedback to train AI speech models. It is a large global company that connects independent contractors to AI projects.

Global

  • Evaluate English-to-and-from-Turkish machine translations to enhance AI translation engines.
  • Work remotely on a project-based schedule with flexible hours.
  • Use your native Turkish and strong English language expertise to assess translation quality.

CrowdGen by Appen is a platform that connects independent contractors with AI training data projects. It is a large, global community of language experts contributing to the improvement of machine translation systems.

Global

  • Evaluate English-to-Hungarian and Hungarian-to-English machine translations for accuracy and fluency.
  • Use your native Hungarian and strong English proficiency to identify and correct translation errors.
  • Work independently on a project basis with flexible hours and remote setup.

We are a platform connecting independent contractors with AI training projects, focusing on improving machine translation engines. We are a large AI data company with a global community of language experts, offering project-based opportunities.

US

  • Review, evaluate, and annotate AI-generated content across text, images, audio, and video.
  • Perform quality checks to ensure accuracy, consistency, and compliance with project guidelines.
  • Identify edge cases and inconsistencies, contribute to high-quality dataset development, and participate in calibration activities.

Welo Data, part of Welocalize, is a global AI data company with over 500,000 contributors that provides high-quality, ethical data for training advanced AI systems. The company supports a diverse, global community across 100+ countries and offers project-based freelance opportunities with flexibility and growth potential.

Germany

  • Evaluate English-to-German and German-to-English machine translations to improve accuracy.
  • Work as an independent contractor with flexible hours on a project basis.
  • Must be a native German speaker with strong English and language expertise.

Appen provides AI training data and machine learning services to improve technology. They operate with a global network of independent contractors, offering flexible project-based work.

Netherlands

  • Label and evaluate photos, graphics, videos, stickers, and designs in Dutch for linguistic accuracy and cultural appropriateness.
  • Assess AI-generated content against Canva's quality bar for Dutch users to shape localised AI experiences.
  • Build and contribute to Dutch-specific datasets and deliver labelled assets on time across varied task types.

Canva is redefining how the world experiences design, empowering users to create visual content. The company has a global team and supports flexible, remote-friendly work, with a focus on collaboration and innovation.

$5–$5/hr
Global

  • Perform data collection, evaluation, and annotation for AI training.
  • Conduct pairwise comparisons and counting tasks.
  • Tag and label objects across audio, video, images, or collected data.

RWS provides AI training data services. They are a global company with a focus on diversity, equity, and inclusion, offering flexible remote work opportunities.

$11–$11/hr
Global

  • Collect, evaluate, and annotate diverse data to improve AI-generated content in Italian.
  • Perform pairwise comparisons, counting tasks, and object tagging across audio, video, images, and text.
  • Work remotely on a flexible, part-time schedule with a long-term contract.

RWS provides technology-enabled language, content management, and intellectual property services. It is a large global company that values diversity and equal opportunity, offering flexible remote work.

Global

  • Create realistic, domain-specific tasks in Spanish/English reflecting Mexico-specific finance practices.
  • Adapt and apply clear scoring rubrics to evaluate AI-generated and human responses.
  • Review and score submissions for accuracy, regulatory alignment, and professional quality.

Lilt's mission is to make the world's information available to everyone, no matter the language they speak. As a global community of linguists and experts, they deliver multilingual AI and human-verified services to Enterprises, Governments, and AI Developers worldwide.