Source Job

18 jobs similar to AI Trainer - Neuroscience Specialist

Jobs ranked by similarity.

$80–$150/hr
UK

  • Review AI-generated responses to clinical scenarios for accuracy and safety.
  • Compare and justify the best responses among multiple model answers.
  • Write improved exemplars and structured feedback to enhance AI model learning.

Prolific is building the largest pool of quality human data in the world, used by over 35,000 AI developers and researchers. It is a platform that connects researchers with a global participant pool for ethically sourced human data.

UK

  • Compare and rank AI-generated responses for accuracy, logic, and safety.
  • Review CS research papers alongside AI summaries to ensure scientific integrity.
  • Fact-check technical data and code for logical flaws and inaccuracies.

Prolific is building the largest pool of quality human data in the world, serving over 35,000 AI developers and researchers. They connect researchers with paid participants to gather high-quality, ethically sourced behavioral data for AI development.

US

  • Evaluate financial documents and reports to verify accuracy and provide AI training data.
  • Respond to AI prompts using financial expertise to teach models complex fiscal concepts.
  • Validate AI outputs against professional financial standards and provide expert feedback.

Prolific builds the world's largest pool of quality human data for AI training. Over 35,000 AI developers and researchers use Prolific, and the company focuses on ethically sourced, diverse human behavioral data.

Global

  • Evaluate LLM responses for accuracy, clarity, and completeness.
  • Fact-check technical claims using authoritative references.
  • Validate code and outputs, and annotate model performance.

Prolific builds the largest pool of high-quality human data for AI development, serving over 35,000 AI developers, researchers, and organizations. They connect researchers with a global community to collect ethically sourced behavioral data.

US

  • Evaluate AI-generated research on environmental topics for scientific accuracy and nuance.
  • Fact-check environmental claims, emissions factors, and legislative requirements.
  • Assess sustainability logic and annotate geospatial data to improve model outputs.

Prolific builds the world's largest pool of quality human data for AI development. With over 35,000 AI developers, researchers, and organizations using the platform, they focus on ethically sourced human behavioral data to train and evaluate AI models.

Canada

  • Apply deep subject-matter expertise to AI model evaluation and large language model projects.
  • Develop challenging domain-specific problems and assess AI responses for accuracy and reasoning.
  • Collaborate with AI research teams to improve training datasets and evaluation methodologies.

Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. It offers a remote, asynchronous work culture and uses AI tools to support recruitment.

  • Create original graduate-to-PhD-level academic problems in your field
  • Write rigorous, step-by-step solutions with exact and verifiable answers
  • Review AI-generated responses to identify specific reasoning errors

Terac is building the world's largest pool of vetted human experts for AI. Researchers, AI labs, and product teams use Terac to recruit, screen, and pay study participants across industries, languages, and skill sets.

Global

  • Evaluate LLM architecture logic for technical accuracy and audit ML code and notebooks for efficiency.
  • Refine RLHF frameworks to align models with human intent and analyze model reasoning in complex chain-of-thought prompts.
  • Benchmark performance by conducting comparative testing between model outputs based on technical metrics.

Prolific connects researchers with a global pool of participants for collecting high-quality human data to train AI models. With over 35,000 users, they focus on ethical data gathering to advance AI capabilities.

Global

  • Create realistic, domain-specific tasks in the target language reflecting local natural sciences practices.
  • Adapt and apply clear scoring rubrics to evaluate AI-generated and human responses.
  • Review and score submissions for accuracy, regulatory alignment, and professional quality.

LILT provides multilingual AI and human-verified services to Enterprises, Governments, and AI Developers. They have a global community of linguists and subject matter experts focused on innovation and excellence.

Global

  • Complete AI training tasks such as analyzing, editing, and writing in Korean.
  • Evaluate and judge AI performance on Korean prompts.
  • Help improve cutting-edge AI models with your expertise.

Prolific builds the largest pool of quality human data for AI development. Over 35,000 developers and researchers use Prolific to gather diverse data from paid participants.

US

  • Evaluate AI-generated wireframes, user flows, and product requirements for usability and logical consistency.
  • Audit accessibility and inclusion, ensuring adherence to modern standards and inclusive design practices.
  • Fact-check UX principles and validate model suggestions against established design patterns and psychological principles.

Prolific builds the world's largest pool of quality human data for AI training. Over 35,000 AI developers and researchers use the platform, and the company fosters a flexible, remote-first culture.

Canada

  • Evaluate AI-generated documents and presentations against quality standards.
  • Apply humanities expertise to identify inaccuracies and cultural issues.
  • Provide structured feedback to improve AI model performance.

A partner company is seeking a humanities evaluator to assess AI-generated content for accuracy and quality. The company emphasizes cultural awareness and critical thinking in a remote, asynchronous work environment.

India

  • Evaluate AI-generated content against domain-specific quality rubrics in humanities, arts, and culture.
  • Review documents, spreadsheets, and presentations for accuracy, relevance, clarity, and overall quality.
  • Provide structured feedback and collaborate with AI research teams to improve model outputs.

A partner company is seeking subject-matter experts to evaluate AI-generated content across humanities, arts, and culture. The company offers a flexible, remote contract environment, with no details on team size provided.

Canada

  • Evaluate AI-generated documents, spreadsheets, and presentation decks against quality rubrics.
  • Identify factual, formatting, visual, and structural issues in professional deliverables.
  • Provide clear, structured feedback to enhance AI output quality and consistency.

This partner company specializes in AI training and evaluation, focusing on improving the quality of AI-generated professional content. Operating as a remote and asynchronous team, they value precision, collaboration, and independent work.

Portugal

  • Design advanced biology problems that challenge frontier AI systems in molecular biology, genetics, or computational biology.
  • Create deterministic tasks with exactly one correct answer and submit complete, verified solutions.
  • Use Python and bioinformatics tools to build problems involving experimental reasoning and computational analysis.

We create high-quality STEM training data for frontier AI models used by leading AI labs. We are a team of experts working to improve model reasoning in scientific domains.

  • Evaluate prompts and AI-generated outputs for accuracy, cultural appropriateness, and brand alignment.
  • Review and correct text, analyze multimedia content, and contribute voice recordings.
  • Apply local cultural insight and consistent evaluation guidelines to ensure high-quality AI training.

Lilt provides multilingual AI and human-verified services to enterprises, governments, and AI developers. They foster a global community of linguists and subject matter experts working on cutting-edge AI and language technology.

Global

  • Evaluate model-generated content across multiple modalities including text, images, audio, and video.
  • Apply defined quality rubrics such as factuality, consistency, and aesthetics to assess outputs.
  • Conduct independent research on unfamiliar topics to make well-supported evaluation judgments.

Our client is a global technology company that helps businesses build, train, and manage AI systems. They offer flexible, remote work with variable workload and a focus on high-quality model evaluation.

Global

  • You will rate and assess the performance of AI models based on their output or behavior.
  • You will label elements of content and assign predefined categories to generate training data.
  • You will create prompts, summaries, and evaluate relevance to improve AI system understanding.

Innodata (Nasdaq: INOD) is a global data engineering company that provides data, evaluation frameworks, and human expertise for AI systems. The company has a 36+ year legacy of delivering high-quality data and outstanding outcomes for customers.