Source Job

Brazil

  • Test and evaluate AI chatbots and language models through structured conversations using assigned criteria.
  • Assess AI-generated responses for quality, relevance, safety, and linguistic accuracy.
  • Submit accurate deliverables such as written evaluations, ratings, and audio recordings within required timelines.

Brazilian Portuguese English AI Evaluation Attention To Detail

20 jobs similar to Brazilian Portuguese - AI Model Rater

Jobs ranked by similarity.

Brazil

  • Review and validate AI training data for accuracy and consistency.
  • Provide constructive feedback to improve data quality.
  • Ensure compliance with project guidelines and escalate issues.

They are a partner company working on innovative AI data projects. The team is global and focused on improving AI accuracy through quality assurance.

Brazil

  • Develop advanced, open-ended questions grounded in realistic scenarios related to sports, equipment, clubs, and events in Brazil.
  • Provide in-depth, accurate answers supported by strong domain knowledge and Brazil-specific context.
  • Follow detailed project guidelines, meet timelines, and maintain high writing quality in Brazilian Portuguese.

RWS is a global leader in language services and AI training, helping shape the future of AI systems. The company employs a diverse, remote workforce and emphasizes quality and domain expertise.

US

  • Evaluate AI-generated content for quality, accuracy, and cultural relevance
  • Apply Castilian Spanish expertise to assess response appropriateness for Spain
  • Provide structured feedback and document decisions to improve AI performance

Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. They use objective, data-driven recruitment processes and prioritize privacy and fairness.

  • Evaluate prompts and AI-generated outputs for accuracy, cultural appropriateness, and brand alignment.
  • Review and correct text, analyze multimedia content, and contribute voice recordings.
  • Apply local cultural insight and consistent evaluation guidelines to ensure high-quality AI training.

Lilt provides multilingual AI and human-verified services to enterprises, governments, and AI developers. They foster a global community of linguists and subject matter experts working on cutting-edge AI and language technology.

US

  • Engage in conversations with a real-time speech-to-speech AI model
  • Evaluate performance on speech recognition, audio quality, conversation flow, and content accuracy
  • Provide accurate ratings based on project guidelines within specified timelines

Appen is a global leader in AI training data and crowd-sourced solutions. They work with a large community of independent contractors to improve AI systems through human evaluation.

US

  • Perform side-by-side comparisons of AI-generated responses.
  • Evaluate outputs for accuracy, relevance, clarity, instruction-following, and overall quality.
  • Assess AI responses across general-purpose questions and answers, web search results, file-based and image-based responses, content-generation tasks, and single-turn and multi-turn conversations.

We are a technology solutions firm headquartered in Bellevue, Washington, with teams across the United States. We help organizations turn complex challenges into meaningful outcomes by connecting strategy and execution across AI, cloud, data, product development, and emerging technology.

Global

  • Evaluate AI-generated text and audio in Catalan for accuracy and natural flow.
  • Provide corrections and constructive feedback on grammar, tone, and cultural context.
  • Complete approximately 10 hours of asynchronous tasks each week via our online platform.

Terac is building the world's largest pool of vetted human experts for AI. Researchers, AI labs, and product teams use Terac to recruit, screen, and pay study participants across industries, languages, and skill sets.

Global

  • Complete AI training tasks such as analyzing, editing, and writing in Korean.
  • Evaluate and judge AI performance on Korean prompts.
  • Help improve cutting-edge AI models with your expertise.

Prolific builds the largest pool of quality human data for AI development. Over 35,000 developers and researchers use Prolific to gather diverse data from paid participants.

Global

  • Provide native-level Canadian French language vetting and QA for AI data projects.
  • Annotate and review AI outputs for grammatical accuracy, cultural context, and naturalness.
  • Develop educational resources and feedback documentation to improve AI alignment.

We are an AI training company that focuses on language alignment and data annotation for AI systems. Our remote team values linguistic precision and cultural nuance.

Global

  • Evaluate prompts and AI-generated outputs for accuracy, clarity, and cultural appropriateness.
  • Review and correct text, analyze multimedia content, and contribute voice recordings.
  • Apply local insight into tone, symbolism, visual cues, and market fit to deliver culturally relevant content.

LILT is an AI company that makes the world's information available to everyone, no matter the language they speak. They work with a global community of linguists and subject matter experts to deliver multilingual AI and human-verified services to Enterprises, Governments, and AI Developers.

Global

  • Side-by-side evaluation of text and voice snippets to assess quality and authenticity.
  • Naturalness assessment of AI-generated audio to rate how natural the voice sounds.
  • Quality control to identify where the AI's tone or pronunciation feels unnatural or culturally mismatched.

Prolific builds the largest pool of quality human data for AI development. Over 35,000 AI developers, researchers, and organizations use Prolific to gather data from paid study participants.

US

  • Review, evaluate, and annotate AI-generated content across text, images, audio, and video.
  • Perform quality checks to ensure accuracy, consistency, and compliance with project guidelines.
  • Identify edge cases and inconsistencies, contribute to high-quality dataset development, and participate in calibration activities.

Welo Data, part of Welocalize, is a global AI data company with over 500,000 contributors that provides high-quality, ethical data for training advanced AI systems. The company supports a diverse, global community across 100+ countries and offers project-based freelance opportunities with flexibility and growth potential.

Netherlands

  • Label and evaluate photos, graphics, videos, stickers, and designs in Dutch for linguistic accuracy and cultural appropriateness.
  • Assess AI-generated content against Canva's quality bar for Dutch users to shape localised AI experiences.
  • Build and contribute to Dutch-specific datasets and deliver labelled assets on time across varied task types.

Canva is redefining how the world experiences design, empowering users to create visual content. The company has a global team and supports flexible, remote-friendly work, with a focus on collaboration and innovation.

Brazil Unlimited PTO

  • Design, build, and deploy AI agents and automations to enhance pre-sales and go-to-market processes.
  • Analyze end-to-end pre-sales operations to identify friction points and implement scalable optimization solutions.
  • Integrate AI tools with CRM, proposal management, and knowledge base platforms to accelerate customer engagement.

The partner company operates at the intersection of AI, automation, and commercial strategy, transforming pre-sales operations. It is a fast-paced, innovative environment focused on experimentation, offering significant ownership and the chance to shape go-to-market operations through emerging technologies.

Spain

  • Evaluate machine translations from English to Spanish (Spain) to improve translation engine accuracy.
  • Work as an independent contractor with flexible hours, completing files at your own pace.
  • Must be a native Spanish speaker with strong English proficiency and reside in Spain.

CrowdGen by Appen connects independent contractors with AI training projects, including machine translation evaluation. The company is part of Appen, a global leader in AI training data, and fosters a flexible, remote work culture.

  • Dive into a cutting-edge AI benchmarking project focused on highly specific professional domains like software engineering, healthcare, and finance.
  • Design realistic scenarios in English or Korean, adapt evaluation rubrics, and review AI responses for accuracy and cultural appropriateness.
  • Contribute to gold-standard solutions that reflect best practices across your target locale and domain.

Lilt is a company that makes the world's information available to everyone, regardless of language, through multilingual AI and human-verified services for Enterprises, Governments, and AI Developers. The company has a global community of linguists and subject matter experts who thrive on innovation and excellence.

Global

  • Contribute to building smarter, more accurate AI by annotating, evaluating, and creating prompts for language models.
  • Work flexibly on your own terms with remote projects that fit your schedule and skills.
  • Be part of a global community of linguists and tech enthusiasts shaping the future of AI.

We are a global AI data company with 500,000+ contributors delivering high-quality, ethical data to train the world's most advanced AI systems. We are building a diverse community in 100+ countries, offering flexible remote work and limitless opportunities for growth.

$35–$35/hr
Germany UK

  • Evaluate AI-generated responses for accuracy, grammar, and cultural relevance in German.
  • Create natural prompts and responses in German to improve conversational datasets.
  • Collaborate with global teams to help refine AI language models.

Welo Data, part of Welocalize, is a global AI data company that provides high-quality, ethical data to train advanced AI systems. With over 500,000 contributors worldwide, they focus on building smarter, more human AI through a diverse, global community.

Canada

  • Evaluate AI-generated documents and presentations against quality standards.
  • Apply humanities expertise to identify inaccuracies and cultural issues.
  • Provide structured feedback to improve AI model performance.

A partner company is seeking a humanities evaluator to assess AI-generated content for accuracy and quality. The company emphasizes cultural awareness and critical thinking in a remote, asynchronous work environment.

Canada

  • Conduct red-team evaluations to identify jailbreaks, prompt injections, and misuse scenarios in conversational AI models.
  • Develop creative adversarial prompts and scenarios to systematically probe model behavior and uncover weaknesses.
  • Generate high-quality human evaluation data by annotating failures and classifying vulnerabilities.

The partner company is a technology organization focused on AI safety and responsible AI development. They work with a remote, asynchronous team to improve the robustness of conversational AI systems.