Source Job

20 jobs similar to Speech AI Evaluation Specialist - Chinese Simplified (Malaysia)

Jobs ranked by similarity.

US

  • Review, evaluate, and annotate AI-generated content across text, images, audio, and video.
  • Perform quality checks to ensure accuracy, consistency, and compliance with project guidelines.
  • Identify edge cases and inconsistencies, contribute to high-quality dataset development, and participate in calibration activities.

Welo Data, part of Welocalize, is a global AI data company with over 500,000 contributors that provides high-quality, ethical data for training advanced AI systems. The company supports a diverse, global community across 100+ countries and offers project-based freelance opportunities with flexibility and growth potential.

US

  • Engage in conversations with a real-time speech-to-speech AI model
  • Evaluate performance on speech recognition, audio quality, conversation flow, and content accuracy
  • Provide accurate ratings based on project guidelines within specified timelines

Appen is a global leader in AI training data and crowd-sourced solutions. They work with a large community of independent contractors to improve AI systems through human evaluation.

Global

  • Evaluate AI-generated text and audio in Catalan for accuracy and natural flow.
  • Provide corrections and constructive feedback on grammar, tone, and cultural context.
  • Complete approximately 10 hours of asynchronous tasks each week via our online platform.

Terac is building the world's largest pool of vetted human experts for AI. Researchers, AI labs, and product teams use Terac to recruit, screen, and pay study participants across industries, languages, and skill sets.

Global

  • Evaluate AI-generated text and voice snippets in Marathi for quality and authenticity.
  • Listen to audio clips and rate how natural the AI voice sounds.
  • Provide feedback on tone, pronunciation, and cultural context.

Prolific is an AI data platform that connects researchers with a global pool of participants to gather high-quality, ethically sourced human data. With over 35,000 AI developers and organizations using the platform, Prolific is building the largest pool of quality human data to train AI models.

Global

  • Contribute to building smarter, more accurate AI by annotating, evaluating, and creating prompts for language models.
  • Work flexibly on your own terms with remote projects that fit your schedule and skills.
  • Be part of a global community of linguists and tech enthusiasts shaping the future of AI.

We are a global AI data company with 500,000+ contributors delivering high-quality, ethical data to train the world's most advanced AI systems. We are building a diverse community in 100+ countries, offering flexible remote work and limitless opportunities for growth.

Global

  • Side-by-side evaluation of text and voice snippets to assess quality and authenticity.
  • Naturalness assessment of AI-generated audio to rate how natural the voice sounds.
  • Quality control to identify where the AI's tone or pronunciation feels unnatural or culturally mismatched.

Prolific builds the largest pool of quality human data for AI development. Over 35,000 AI developers, researchers, and organizations use Prolific to gather data from paid study participants.

US Canada UK

  • Evaluate AI-generated responses in Dutch for language quality, customer experience, technical accuracy, and JSON structure.
  • Participate in real-time conversational scenarios with AI systems and review transcripts.
  • Capture session artifacts and provide written justifications using structured rubrics.

An enterprise client is hiring contract experts to evaluate AI-generated audio outputs for quality and technical accuracy. The project offers flexible scheduling and remote work, with a focus on language fluency and coding skills.

Global

  • Have live, natural conversations in Japanese with other participants over video call.
  • Discuss everyday topics without scripts, letting the conversation flow naturally.
  • Help train AI models to understand natural speech patterns, tone, and cultural nuance.

Prolific is building the world's largest pool of quality human data, used by over 35,000 AI developers and researchers. They believe in integrating diverse human perspectives to improve AI systems.

US

  • Evaluate AI-generated content for quality, accuracy, and cultural relevance
  • Apply Castilian Spanish expertise to assess response appropriateness for Spain
  • Provide structured feedback and document decisions to improve AI performance

Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. They use objective, data-driven recruitment processes and prioritize privacy and fairness.

Brazil

  • Test and evaluate AI chatbots and language models through structured conversations using assigned criteria.
  • Assess AI-generated responses for quality, relevance, safety, and linguistic accuracy.
  • Submit accurate deliverables such as written evaluations, ratings, and audio recordings within required timelines.

This company specializes in AI development and data evaluation, focusing on improving generative AI systems. The organization operates with a flexible, project-based team and values linguistic expertise.

Global

  • Evaluate prompts and AI-generated outputs for accuracy, clarity, and cultural appropriateness.
  • Review and correct text, analyze multimedia content, and contribute voice recordings.
  • Apply local insight into tone, symbolism, visual cues, and market fit to deliver culturally relevant content.

LILT is an AI company that makes the world's information available to everyone, no matter the language they speak. They work with a global community of linguists and subject matter experts to deliver multilingual AI and human-verified services to Enterprises, Governments, and AI Developers.

  • Evaluate prompts and AI-generated outputs for accuracy, cultural appropriateness, and brand alignment.
  • Review and correct text, analyze multimedia content, and contribute voice recordings.
  • Apply local cultural insight and consistent evaluation guidelines to ensure high-quality AI training.

Lilt provides multilingual AI and human-verified services to enterprises, governments, and AI developers. They foster a global community of linguists and subject matter experts working on cutting-edge AI and language technology.

India

  • Label, annotate, and evaluate Hindi content for linguistic quality, accuracy, and cultural appropriateness.
  • Build and contribute to Hindi-specific datasets to support internationalisation of AI features.
  • Review and refine labels based on feedback to maintain consistency across task types.

Canva is a design platform that redefines how the world experiences design. We have a global team that supports remote collaboration and values diverse skills and backgrounds.

Global

  • Engage in live video conversations in Bengali to help train AI models.
  • Discuss everyday topics naturally with other participants, with no scripts.
  • Choose your own schedule and work remotely with flexible task-based hours.

Prolific builds a platform connecting researchers with a global pool of participants to collect high-quality human data for AI development. With over 35,000 organizations using the platform, they focus on ethical data collection and flexible, remote participation.

Netherlands

  • Label and evaluate photos, graphics, videos, stickers, and designs in Dutch for linguistic accuracy and cultural appropriateness.
  • Assess AI-generated content against Canva's quality bar for Dutch users to shape localised AI experiences.
  • Build and contribute to Dutch-specific datasets and deliver labelled assets on time across varied task types.

Canva is redefining how the world experiences design, empowering users to create visual content. The company has a global team and supports flexible, remote-friendly work, with a focus on collaboration and innovation.

Global

  • Serve as the primary Hindi language expert, owning quality, data, and product performance.
  • Coordinate with the Human Evals Manager to manage annotation quality and deliver feedback.
  • Work independently to audit datasets, investigate issues, and develop onboarding materials.

Cartesia architects AI that learns from and interacts with the world, pioneering state space models for efficient large-scale foundations. Backed by leading investors and staffed by Stanford AI Lab PhDs, it fosters a fast-paced, inclusive culture.

Global

  • Evaluate English-to-and-from-Turkish machine translations to enhance AI translation engines.
  • Work remotely on a project-based schedule with flexible hours.
  • Use your native Turkish and strong English language expertise to assess translation quality.

CrowdGen by Appen is a platform that connects independent contractors with AI training data projects. It is a large, global community of language experts contributing to the improvement of machine translation systems.

Brazil

  • Review and validate AI training data for accuracy and consistency.
  • Provide constructive feedback to improve data quality.
  • Ensure compliance with project guidelines and escalate issues.

They are a partner company working on innovative AI data projects. The team is global and focused on improving AI accuracy through quality assurance.

Brazil

  • Review and analyze short-form video content to identify all languages in spoken audio and on-screen text.
  • Validate that language and cultural elements accurately reflect Brazilian Portuguese for machine learning datasets.
  • Perform quality checks and apply classification guidelines to generate accurate ground-truth labels.

Jobgether is an AI-powered job matching platform that connects candidates with partner companies. They use technology to ensure fast, objective candidate reviews and partner with firms in international projects.