Remote Data Jobs · Linguistics

Job listings

  • Perform linguistic tasks including translation, transcription, annotation, and evaluation for AI models.
  • Review and validate audio, video, and text data to ensure accuracy and consistency.
  • Apply strong Marathi and English language skills to produce high-quality linguistic data.

Sigma AI India works on improving AI systems through high-quality linguistic data, focusing on translation, annotation, and content creation. They are a partner company offering freelance opportunities in a fully remote environment with global AI projects.

  • Conduct live, natural video conversations in Vietnamese with other participants to train AI models.
  • Discuss everyday topics without scripts, allowing conversations to flow naturally for authentic speech data.
  • Work on a task-based, flexible schedule with no minimum hours, choosing tasks based on details and pay.

Prolific is building the world's largest pool of quality human data for AI developers, researchers, and organizations. With over 35,000 developers using their platform, they offer a flexible, ethical, and remote-friendly culture focused on integrating diverse human perspectives into AI development.

  • Evaluate AI-generated text and voice snippets in Marathi for quality and authenticity.
  • Listen to audio clips and rate how natural the AI voice sounds.
  • Provide feedback on tone, pronunciation, and cultural context.

Prolific is an AI data platform that connects researchers with a global pool of participants to gather high-quality, ethically sourced human data. With over 35,000 AI developers and organizations using the platform, Prolific is building the largest pool of quality human data to train AI models.

  • Have live, natural conversations in Japanese with other participants over video call.
  • Discuss everyday topics without scripts, letting the conversation flow naturally.
  • Help train AI models to understand natural speech patterns, tone, and cultural nuance.

Prolific is building the world's largest pool of quality human data, used by over 35,000 AI developers and researchers. They believe in integrating diverse human perspectives to improve AI systems.

$15–$15/hr

  • Rating and assessing the performance of AI models based on their output or behavior.
  • Labeling and categorizing content to train machine learning models.
  • Generating prompts, responses, and summaries to improve language model reasoning.

Innodata is a global data engineering company that provides data, evaluation frameworks, and human expertise for AI systems. With over 36 years of experience, the company focuses on enabling responsible AI advancement.

  • Serve as the primary Hindi language expert, owning quality, data, and product performance.
  • Coordinate with the Human Evals Manager to manage annotation quality and deliver feedback.
  • Work independently to audit datasets, investigate issues, and develop onboarding materials.

Cartesia architects AI that learns from and interacts with the world, pioneering state space models for efficient large-scale foundations. Backed by leading investors and staffed by Stanford AI Lab PhDs, it fosters a fast-paced, inclusive culture.