Design realistic Manufacturing and Industrial Operations scenarios in Punjabi or English, reflecting Indian operational contexts.
Review AI and human-generated responses for factual accuracy, operational efficiency, and industrial realism.
Contribute to 'gold standard' solutions for production, procurement, and supply chain challenges.
LILT provides multilingual AI and human-verified services to enterprises, governments, and AI developers. It fosters a global community of linguists and subject matter experts, prioritizing innovation, excellence, and flexible collaboration.
Design realistic Retail Trade scenarios in Korean or English grounded in Korean operational contexts.
Adapt structured evaluation rubrics for operational problem-solving and customer interactions.
Review AI and human-generated responses for factual accuracy and operational realism.
LILT is a company that provides AI and human-verified multilingual services to Enterprises, Governments, and AI Developers. They have a global community of linguists and subject matter experts, and they emphasize innovation and excellence.
Create and review realistic professional services scenarios in Nepali or English for AI benchmarking in Indian corporate contexts.
Adapt evaluation rubrics for analytical reasoning, technical problem-solving, and project coordination tasks.
Review AI and human-generated responses for factual accuracy, professional standards, and operational realism.
LILT provides multilingual AI and human-verified services to enterprises and governments worldwide. The company fosters a global, innovative community of linguists and subject matter experts dedicated to advancing human knowledge.
Design realistic scenarios in your target language or English grounded in operational contexts.
Adapt structured evaluation rubrics and review AI/human responses for accuracy, quality, and cultural appropriateness.
Contribute to gold-standard solutions reflecting best practices across target locale and domain.
LILT is an AI language company that provides multilingual AI and human-verified services to enterprises, governments, and AI developers. It operates with a global community of linguists and subject matter experts focused on innovation and excellence.
Design and engineer challenging benchmark tasks for evaluating coding agents in multilingual terminal environments.
Create authentic task environments using native language assets and identify model failure points.
Participate in rigorous quality assurance processes including calibration and audit of benchmark tasks.
The hiring company specializes in AI evaluation and multilingual language technology. They are a global team of engineers and linguists working on cutting-edge AI systems.
Design and build rigorous, verifiable Terminal-Bench tasks that test multilingual robustness in LLMs across prompt language effects and encoding edge cases.
Create realistic task environments with datasets and files in your native language, ensuring assets remain in the target language to genuinely measure multilingual handling.
Calibrate task difficulty by analyzing execution logs and participate in a 4-layer human quality control process to ensure benchmark integrity.
LILT is an AI and language technology company whose mission is to make the world's information available to everyone, regardless of language. They operate with a global community of linguists, engineers, and subject matter experts, fostering a culture of innovation and excellence.
Contribute to shaping safer, smarter AI by joining a global network of linguists and culturally aware contributors.
Work on flexible, remote projects in annotation, evaluation, and prompt creation, always on your terms.
Get first access to projects that match your skills, from short tasks to multi-week assignments.
Welo Data, part of Welocalize, is a global AI data company with a network of over 500,000 contributors. They build smarter, more human AI by offering flexible, remote projects to a diverse community in over 100 countries, emphasizing growth and work-life balance.
Evaluate and score AI-generated customer interactions across Foundational, Experiential, and Operational dimensions using complex rubrics.
Benchmark informational and transactional customer queries against authoritative business sources to ensure accuracy.
Participate in dual-review processes and daily calibration audits to maintain inter-rater agreement and quality standards.
Innodata is a global data engineering company that enables the responsible advancement of artificial intelligence by providing data, evaluation frameworks, and human expertise. The company has a 36+ year legacy delivering high-quality data and outstanding outcomes for customers.
Create and review realistic Media and Information Services scenarios in Gujarati or English for AI benchmarking.
Adapt evaluation rubrics and contribute to gold-standard solutions for editorial, production, and content challenges.
Apply 5+ years of professional experience in media, broadcasting, journalism, or content production.
Lilt provides multilingual AI and human-verified services to enterprises, governments, and AI developers worldwide. It operates as a global community of linguists and subject matter experts, emphasizing innovation and excellence.
Review and evaluate annotation and audit work according to project guidelines.
Perform quality checks across Simplified and Traditional Chinese content.
Identify errors, inconsistencies, and recurring quality issues.
Welo Data provides AI services including data annotation and quality assurance for generative AI systems. They are a global community of AI specialists working on next-generation AI projects.
Work from anywhere in the world with a flexible schedule that fits your lifestyle.
Access a variety of projects in AI/Data Annotation, Customer Support, and Back-office Operations.
Be the first to receive offers for new upcoming projects as a priority talent pool member.
We are an international company providing data annotation services for AI, as well as customer support and back-office solutions to clients worldwide. We are building a talent pool of individuals seeking ongoing access to freelance opportunities, offering flexibility and diverse projects.
Evaluate AI-generated work products in real estate, hospitality, and events using quality rubrics.
Identify factual, aesthetic, and presentation errors and provide actionable feedback.
Apply industry expertise to distinguish realistic, commercially sound work from generic AI content.
The company develops AI systems and evaluates their outputs for quality. They seek experienced industry professionals for flexible remote contract work.
Evaluate prompts and AI-generated outputs for accuracy, clarity, and cultural appropriateness.
Review and correct text, analyze multimedia content, and contribute voice recordings.
Apply careful judgment to ensure high-quality results aligned with task objectives.
LILT provides multilingual AI and human-verified services to enterprises, governments, and AI developers. The company has a global community of linguists and language professionals committed to innovation and excellence.