Source Job

$100–$150/hr
Canada

  • Evaluate AI-generated slides, spreadsheets, and documents for real-world usability and professional quality.
  • Assess outputs for accuracy, clarity, relevance, and alignment with data science standards.
  • Provide structured written feedback to help improve AI systems and their outputs.

Data Science Microsoft Office Google Workspace Analytical Skills Written Communication

17 jobs similar to Data Science Expert

Jobs ranked by similarity.

India

  • Evaluate AI-generated documents, spreadsheets, and presentation decks for accuracy and professional quality.
  • Assess visual and aesthetic quality including layout, formatting, and readability.
  • Provide clear, structured written feedback to identify issues and improve AI outputs.

Our partner is a company focused on improving AI systems through quality evaluation. They offer a flexible, remote work environment for independent contractors.

India

  • Analyze, evaluate, and review diverse datasets to support AI system training and improvement.
  • Assess AI-generated content for accuracy, relevance, consistency, and quality, providing actionable feedback.
  • Work independently in a remote, digital-first environment, managing multiple tasks and deadlines.

US

  • Evaluate AI-generated market research and competitive intelligence artifacts for accuracy, rigor, and quality.
  • Apply structured rubrics to assess deliverables and identify factual inaccuracies and analytical gaps.
  • Provide clear, actionable written feedback to support evaluation decisions and improve AI training.

The partner company specializes in evaluating AI-generated market research and competitive intelligence content. They hire independent contractors for flexible remote engagements.

US

  • Review and advise on evaluation criteria and scoring rubrics for AI-generated outputs.
  • Create, edit, and validate high-quality benchmark tasks and reference data for AI training.
  • Analyze model failures, including hallucinations and flawed reasoning, providing expert explanations.

The company specializes in AI evaluation and training, helping define standards for next-generation AI systems. They operate as a partner company that manages applications and next steps, valuing autonomy and expertise in their consultants.

US

  • Evaluate AI-generated work products in real estate, hospitality, and events using quality rubrics.
  • Identify factual, aesthetic, and presentation errors and provide actionable feedback.
  • Apply industry expertise to distinguish realistic, commercially sound work from generic AI content.

The company develops AI systems and evaluates their outputs for quality. They seek experienced industry professionals for flexible remote contract work.

US

  • Evaluate AI-generated documents, spreadsheets, and presentations against domain-specific quality standards.
  • Assess outputs for factual accuracy, procurement relevance, completeness, clarity, and practical applicability.
  • Provide structured written feedback identifying issues and opportunities for improvement.

This partner company specializes in evaluating AI-generated work products for public-sector procurement and RFI responses. The company size and culture are not specified in the posting, but the engagement is flexible and fully remote.

Australia

  • Evaluate AI-generated responses for relevance, accuracy, and personalization using personalized prompts and data from connected Google applications.
  • Identify subtle issues such as incorrect assumptions, irrelevant recommendations, inconsistencies, and inappropriate personalization.
  • Provide clear, detailed, and structured feedback to support improvements to AI models and personalization systems.

Our partner company is seeking an AI Response Quality Evaluator to improve AI-generated responses. This is a project-based contract role with a remote, independent working environment and a duration of up to 16 weeks.

USA

  • Conduct structured quality audits on spatial datasets using mapping platforms and GIS-standard evaluation tools.
  • Evaluate and cross-reference complex geographic and business data against multiple authoritative sources.
  • Perform root-cause analysis and document discrepancies to refine mapping algorithms and improve geographic intelligence.

TELUS Digital helps build better AI models by leveraging a global community of over one million contributors who collect, enhance, and validate content. The company fosters a diverse, remote-first culture focused on intellectual rigor and structured data verification.

US

  • Apply expertise in incident management and SRE to evaluate AI-generated documents, spreadsheets, and slide decks for technical accuracy and operational rigor.
  • Assess outputs against real-world reliability practices, identifying factual, technical, and reasoning errors.
  • Provide clear, structured written feedback and collaborate asynchronously with a research team to refine evaluation approaches.

This partner company focuses on AI evaluation and development, seeking experienced professionals to assess AI-generated work products. They offer flexible remote work and independent contractor engagements with weekly payments.

India

  • Evaluate AI-generated content across humanities, arts, and culture domains for accuracy and quality.
  • Provide structured feedback and identify issues in AI outputs.
  • Work remotely on a flexible schedule contributing to AI system improvement.

The company specializes in developing and improving AI systems through expert human evaluation. They operate as a remote, flexible organization that values specialized domain knowledge and critical analysis.

UK 5w PTO

  • Quality-check ocean freight contracts and rates against source documents.
  • Collaborate with international vendor partners to resolve discrepancies.
  • Use AI tools to streamline QA workflows and improve efficiency.

The company specializes in ocean freight data quality assurance. It is a remote-first, international team with a feedback-driven culture that values ownership and continuous improvement.

Global

  • Create and review realistic professional services scenarios in Nepali or English for AI benchmarking in Indian corporate contexts.
  • Adapt evaluation rubrics for analytical reasoning, technical problem-solving, and project coordination tasks.
  • Review AI and human-generated responses for factual accuracy, professional standards, and operational realism.

LILT provides multilingual AI and human-verified services to enterprises and governments worldwide. The company fosters a global, innovative community of linguists and subject matter experts dedicated to advancing human knowledge.

India

  • Evaluate AI-generated financial documents, spreadsheets, and presentations for accuracy and quality.
  • Provide clear, structured written feedback to identify issues and suggest improvements.
  • Use your FP&A and corporate finance expertise to train and enhance AI systems.

Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. The company operates remotely and collaborates with partners to find qualified candidates for specialized roles.

Global

  • Assess software engineering tasks for technical accuracy, realism, and reproducibility.
  • Provide actionable feedback on codebase integration issues and logic errors.
  • Ensure AI training workflows are rigorous and practically applicable.

Project World Wide sources experienced technical specialists for AI training task auditing. This freelance contract opportunity focuses on ensuring technical rigor and accuracy in AI workflows.

US

  • Evaluate AI-generated Azerbaijani text for naturalness and authenticity.
  • Compare text snippets and provide quality control on cultural nuance.
  • Rate AI-generated text and tag data on tone and naturalness.

Prolific is building the largest pool of quality human data in the world, with over 35,000 AI developers and researchers using its platform. They connect researchers with paid participants to collect ethically sourced human behavioral data and feedback.

Germany

  • Label, annotate, and evaluate German-language content including photos, graphics, and videos for linguistic and cultural accuracy.
  • Evaluate AI-generated content against Canva's quality bar for German users to shape language experiences.
  • Build and contribute to German-specific datasets to support the internationalization of Canva AI features.

Canva is a design platform redefining how the world experiences design. It is a global company with a large user base, known for its innovative culture and focus on AI-powered features.

$35–$75/hr
Global

  • Evaluate model outputs in humanities fields for factual accuracy, logical coherence, and ideological bias.
  • Create exemplary responses and datasets emphasizing intellectual honesty and thorough source evaluation.
  • Collaborate with engineering teams to design evaluation tasks and define desired model behavior.

SpaceXAI creates AI systems to understand the universe and aid humanity in its pursuit of knowledge. The team is small, highly motivated, and focused on engineering excellence with a flat organizational structure.