Test and evaluate AI chatbots and language models through structured conversations using assigned criteria.
Assess AI-generated responses for quality, relevance, safety, and linguistic accuracy.
Submit accurate deliverables such as written evaluations, ratings, and audio recordings within required timelines.
This company specializes in AI development and data evaluation, focusing on improving generative AI systems. The organization operates with a flexible, project-based team and values linguistic expertise.
Review short-form video clips to identify languages in audio and on-screen text.
Validate media content against your target locale for authenticity and cultural relevance.
Generate accurate ground truth labels using internal classification tools following strict guidelines.
RWS is a technology-enabled language services company that helps global organizations connect with their audiences. With a large team of linguists and AI specialists, RWS fosters an inclusive culture focused on innovation and quality.
Develop advanced, open-ended questions grounded in realistic scenarios related to sports, equipment, clubs, and events in Brazil.
Provide in-depth, accurate answers supported by strong domain knowledge and Brazil-specific context.
Follow detailed project guidelines, meet timelines, and maintain high writing quality in Brazilian Portuguese.
RWS is a global leader in language services and AI training, helping shape the future of AI systems. The company employs a diverse, remote workforce and emphasizes quality and domain expertise.
Review, evaluate, and annotate AI-generated content across text, images, audio, and video.
Perform quality checks to ensure accuracy, consistency, and compliance with project guidelines.
Identify edge cases and inconsistencies, contribute to high-quality dataset development, and participate in calibration activities.
Welo Data, part of Welocalize, is a global AI data company with over 500,000 contributors that provides high-quality, ethical data for training advanced AI systems. The company supports a diverse, global community across 100+ countries and offers project-based freelance opportunities with flexibility and growth potential.
Evaluate prompts and AI-generated outputs for accuracy, cultural appropriateness, and brand alignment.
Review and correct text, analyze multimedia content, and contribute voice recordings.
Apply local cultural insight and consistent evaluation guidelines to ensure high-quality AI training.
Lilt provides multilingual AI and human-verified services to enterprises, governments, and AI developers. They foster a global community of linguists and subject matter experts working on cutting-edge AI and language technology.
Review pre-seeded questions paired with images to provide accurate 'golden' answers.
Carefully examine image content to identify relevant details for each question.
Flag any images that are unclear, corrupted, or insufficient to answer the question.
Welo Data is an AI services company specializing in data annotation and validation. They operate as a remote-friendly organization with a focus on quality and detail-oriented work.
Collect, evaluate, and annotate diverse data to improve AI-generated content in Italian.
Perform pairwise comparisons, counting tasks, and object tagging across audio, video, images, and text.
Work remotely on a flexible, part-time schedule with a long-term contract.
RWS provides technology-enabled language, content management, and intellectual property services. It is a large global company that values diversity and equal opportunity, offering flexible remote work.
Label and evaluate photos, graphics, videos, stickers, and designs in Dutch for linguistic accuracy and cultural appropriateness.
Assess AI-generated content against Canva's quality bar for Dutch users to shape localised AI experiences.
Build and contribute to Dutch-specific datasets and deliver labelled assets on time across varied task types.
Canva is redefining how the world experiences design, empowering users to create visual content. The company has a global team and supports flexible, remote-friendly work, with a focus on collaboration and innovation.
Label, annotate, and evaluate Hindi content for linguistic quality, accuracy, and cultural appropriateness.
Build and contribute to Hindi-specific datasets to support internationalisation of AI features.
Review and refine labels based on feedback to maintain consistency across task types.
Canva is a design platform that redefines how the world experiences design. We have a global team that supports remote collaboration and values diverse skills and backgrounds.
Evaluate AI-generated text and voice snippets in Marathi for quality and authenticity.
Listen to audio clips and rate how natural the AI voice sounds.
Provide feedback on tone, pronunciation, and cultural context.
Prolific is an AI data platform that connects researchers with a global pool of participants to gather high-quality, ethically sourced human data. With over 35,000 AI developers and organizations using the platform, Prolific is building the largest pool of quality human data to train AI models.
Evaluate prompts and AI-generated outputs for accuracy, clarity, and cultural appropriateness.
Review and correct text, analyze multimedia content, and contribute voice recordings.
Apply local insight into tone, symbolism, visual cues, and market fit to deliver culturally relevant content.
LILT is an AI company that makes the world's information available to everyone, no matter the language they speak. They work with a global community of linguists and subject matter experts to deliver multilingual AI and human-verified services to Enterprises, Governments, and AI Developers.
Evaluate and rank model outputs, stress-test models for failure modes, and create high-quality datasets with detailed rubrics.
Annotate and correct multimodal data, maintain consistency through calibration exercises, and adapt to evolving task types.
Report on model performance trends and provide clear feedback to cross-functional partners on model successes and failures.
Cohere is a security-first enterprise AI company that builds cutting-edge foundation models and end-to-end products for real-world business problems. It is a global technology company with offices in Toronto, San Francisco, London, New York, Montreal, Seoul, Germany, and Paris, staffed by a team of passionate researchers, engineers, and designers.
Provide native-level Canadian French language vetting and QA for AI data projects.
Annotate and review AI outputs for grammatical accuracy, cultural context, and naturalness.
Develop educational resources and feedback documentation to improve AI alignment.
We are an AI training company that focuses on language alignment and data annotation for AI systems. Our remote team values linguistic precision and cultural nuance.
Evaluate AI-generated speech in Chinese Simplified for relevance and accuracy.
Participate in short voice conversations with AI models using designated platforms.
Provide objective ratings and feedback based on project guidelines.
RWS is a global language services and technology company that provides AI training and evaluation solutions. They are a large employer with a diverse and inclusive culture, employing thousands worldwide.
Engage in conversations with a real-time speech-to-speech AI model
Evaluate performance on speech recognition, audio quality, conversation flow, and content accuracy
Provide accurate ratings based on project guidelines within specified timelines
Appen is a global leader in AI training data and crowd-sourced solutions. They work with a large community of independent contractors to improve AI systems through human evaluation.
Evaluate AI-generated responses for accuracy, grammar, and cultural relevance in German.
Create natural prompts and responses in German to improve conversational datasets.
Collaborate with global teams to help refine AI language models.
Welo Data, part of Welocalize, is a global AI data company that provides high-quality, ethical data to train advanced AI systems. With over 500,000 contributors worldwide, they focus on building smarter, more human AI through a diverse, global community.
Source legally usable, public PDF documents in your designated language.
Verify that each document meets open-source or public domain licensing requirements.
Upload the collected files to our research platform and provide basic metadata.
Terac is building the world's largest pool of vetted human experts for AI. They recruit, screen, and pay study participants across industries, languages, and skill sets.
Evaluate AI-generated text and audio in Catalan for accuracy and natural flow.
Provide corrections and constructive feedback on grammar, tone, and cultural context.
Complete approximately 10 hours of asynchronous tasks each week via our online platform.
Terac is building the world's largest pool of vetted human experts for AI. Researchers, AI labs, and product teams use Terac to recruit, screen, and pay study participants across industries, languages, and skill sets.
Review search results and evaluate their relevance to user queries
Answer true/false questions about content quality
Rate search results based on guidelines to improve AI systems
Welo Data provides AI services and data validation to improve search engine and AI systems. They are a remote-first company with a focus on quality and support for their contractors.