Evaluate AI-generated Macedonian text for naturalness and cultural authenticity.
Compare text side-by-side to assess quality and nuance.
Provide feedback on tone, register, and word choice to improve AI language models.
Prolific is building the largest pool of quality human data in the world, used by over 35,000 AI developers and researchers. We connect a global community of participants with researchers to ethically source human behavior and feedback for AI development.
Evaluate prompts and AI-generated outputs for accuracy, clarity, and cultural appropriateness.
Review and correct text, analyze multimedia content, and contribute voice recordings.
Apply careful judgment to ensure high-quality results aligned with task objectives.
LILT provides multilingual AI and human-verified services to enterprises, governments, and AI developers. The company has a global community of linguists and language professionals committed to innovation and excellence.
Collect data according to detailed project guidelines and specifications.
Ensure collected data meets the project's required standards and apply feedback for corrections.
Maintain consistent quality and productivity throughout the project.
Our enterprise client is a leading provider of high-quality, diverse datasets for AI model development. They are seeking skilled contributors to support AI training initiatives focused on improving AI-powered language and speech models.
Label, annotate, and evaluate German-language content including photos, graphics, and videos for linguistic and cultural accuracy.
Evaluate AI-generated content against Canva's quality bar for German users to shape language experiences.
Build and contribute to German-specific datasets to support the internationalization of Canva AI features.
Canva is a design platform redefining how the world experiences design. It is a global company with a large user base, known for its innovative culture and focus on AI-powered features.
Evaluate AI-generated Azerbaijani text for naturalness and authenticity.
Compare text snippets and provide quality control on cultural nuance.
Rate AI-generated text and tag data on tone and naturalness.
Prolific is building the largest pool of quality human data in the world, with over 35,000 AI developers and researchers using its platform. They connect researchers with paid participants to collect ethically sourced human behavioral data and feedback.
Review text or media samples based on provided project guidelines
Apply accurate labels and categorizations to diverse data sets
Evaluate AI-generated responses for clarity, safety, and factual accuracy
Terac is building the world's largest pool of vetted human experts for AI. Researchers, AI labs, and product teams use Terac to recruit, screen, and pay study participants across industries, languages, and skill sets.
Evaluate and edit AI-generated translations between German and English for accuracy and fluency.
Translate content between German and English while maintaining cultural relevance and context.
Annotate translation errors and provide feedback to improve AI model performance.
This partner company is seeking a German translator for an AI training project. It offers a flexible, remote contract opportunity for language professionals to improve AI translation quality.
Evaluate Arabic (Egypt) claims for accuracy using reliable sources and evidence.
Research and validate information, assess source credibility, and document findings.
Work independently on a freelance basis to improve AI-generated information quality.
Welo Data, part of Welocalize, is a global AI data company delivering high-quality, ethical data to train advanced AI systems. With 500,000+ contributors in 100+ countries, they offer a flexible, supportive community for remote freelance work.
Record high-quality French speech samples for AI training datasets.
Evaluate AI-generated French audio for pronunciation, tone, and naturalness.
Provide detailed feedback and collaborate with technical teams to refine voice models.
The hiring partner is an innovative technology company focused on training cutting-edge conversational and expressive AI voice systems. They operate with a remote-first, autonomous freelance culture, collaborating with global talent to shape next-generation speech technology.
Review completed Display Transcription tasks for accuracy and consistency.
Verify correct punctuation, capitalization, numbers, dates, and text normalization.
Apply quality standards and meet productivity targets while working independently.
TSMG is a company that specializes in AI language data collection and quality control for speech recognition systems. They offer flexible project-based work and provide training to contributors, with a focus on accuracy and consistency.
Evaluate and label AI model outputs to improve performance and alignment with project guidelines.
Create prompts, rewrite text, and generate training data for large language models.
Work on flexible, remote, project-based tasks while helping shape the future of AI.
Innodata is a global data engineering company that enables the responsible advancement of AI by providing data, evaluation frameworks, and human expertise. With a 36+ year legacy, the company delivers high-quality data and outstanding outcomes for customers.
Design and build rigorous, verifiable Terminal-Bench tasks that test multilingual robustness in LLMs across prompt language effects and encoding edge cases.
Create realistic task environments with datasets and files in your native language, ensuring assets remain in the target language to genuinely measure multilingual handling.
Calibrate task difficulty by analyzing execution logs and participate in a 4-layer human quality control process to ensure benchmark integrity.
LILT is an AI and language technology company whose mission is to make the world's information available to everyone, regardless of language. They operate with a global community of linguists, engineers, and subject matter experts, fostering a culture of innovation and excellence.
Evaluate search results and AI-generated content for quality and relevance.
Conduct online research to verify information and support rating decisions.
Provide feedback and document edge cases to improve AI systems.
TELUS Digital AI is a global AI community of over 1 million contributors helping clients collect, enhance, and train data to build better AI models. They offer flexible remote work and a diverse, inclusive culture.
Evaluate AI-generated Icelandic text for naturalness and authenticity.
Compare side-by-side text snippets to assess quality.
Provide feedback on tone, register, and cultural context.
Prolific is building the largest pool of quality human data in the world, used by over 35,000 AI developers and researchers. They focus on ethically sourced human behavioral data to improve AI systems.
Evaluate search results and AI-generated content for quality, relevance, accuracy, and usefulness.
Conduct online research to verify information and support rating decisions.
Apply rating guidelines consistently and participate in training and calibration sessions.
TELUS Digital AI & Data Solutions partners with a diverse and vibrant community to help our customers enhance their AI and machine learning models. Our global AI community includes over 1 million contributors across 500+ languages and dialects, offering flexible remote and onsite opportunities.
Create realistic, domain-specific mathematics tasks in Arabic reflecting local practices.
Adapt and apply clear scoring rubrics to evaluate AI-generated and human responses.
Review submissions and provide expert feedback for high-quality gold standard solutions.
LILT provides multilingual AI and human-verified services to Enterprises, Governments, and AI Developers worldwide. They have a global community of linguists and subject matter experts who thrive on innovation and excellence.
Design and engineer challenging benchmark tasks for evaluating coding agents in multilingual terminal environments.
Create authentic task environments using native language assets and identify model failure points.
Participate in rigorous quality assurance processes including calibration and audit of benchmark tasks.
The hiring company specializes in AI evaluation and multilingual language technology. They are a global team of engineers and linguists working on cutting-edge AI systems.
Evaluate search results and AI-generated content for quality, relevance, accuracy, and usefulness.
Conduct online research to verify information and support rating decisions.
Participate in training, calibration sessions, and ongoing quality reviews.
TELUS Digital is a global AI community that helps customers collect, enhance, train, translate, and localize content to build better AI models. It consists of a vibrant network of over 1 million contributors from diverse backgrounds.
Comparing text and voice snippets to assess quality and authenticity.
Listening to AI-generated audio and rating how natural the voice sounds.
Identifying where AI tone or pronunciation feels unnatural or culturally mismatched.
Prolific builds the world's largest pool of quality human data, connecting AI developers and researchers with paid study participants. Over 35,000 organizations use our platform for ethically sourced human behavioral data to improve AI.
Record high-quality Finnish audio samples for AI training data with proper tone and emotion.
Evaluate AI-generated Finnish speech for pronunciation, naturalness, and stylistic accuracy.
Collaborate with technical teams to refine voice design standards and provide structured feedback.
This role is part of a global AI training project that improves Finnish speech technology through voice recordings and evaluation. The team values autonomy and flexible collaboration, with a focus on high-quality, natural AI communication.