Remote Data Jobs · Critical Thinking

Job listings

  • Monitor performance indicators (KPIs), identify trends, and propose operational improvements.
  • Develop analyses and management reports, transforming data into strategic information for decision support.
  • Prepare executive presentations and identify process bottlenecks, recommending continuous improvement actions.

QIMA is a global company specializing in inspection, audit, certification and quality control services for various industries and agribusiness. Present in over 100 countries, the company invests in innovation, technology, and employee development, valuing committed and ethical people.

  • Review and analyze geospatial scenarios involving uncertain and complex real-world data.
  • Evaluate AI-generated outputs for accuracy, relevance, and spatial reasoning quality.
  • Collaborate with technical teams to improve evaluation frameworks and model performance.

This role is offered by a partner company leveraging geospatial expertise to train AI systems. The company is a growing organization with a remote-first culture and a focus on flexible, project-based work.

  • Evaluate simulated advertiser-AI conversations for technical accuracy and campaign structure.
  • Fact-check platform strategies against correct hierarchy and full-funnel metrics.
  • Write clear, actionable feedback to correct errors and improve AI model performance.

RWS specializes in AI training data and language services, providing data annotation and evaluation solutions to improve AI model reliability. The company fosters a culture of diversity and inclusion, operating as a global employer with a focus on equal opportunity.

  • Evaluate and assess AI model outputs based on predefined quality, accuracy, relevance, and behavioral guidelines.
  • Annotate, classify, and label text, images, or audio to support AI model training.
  • Create prompts and generate high-quality responses to improve language model reasoning capabilities.

Jobgether uses an AI-powered matching process to connect candidates with hiring companies. They focus on efficient, fair recruitment and handle data privacy in compliance with GDPR.

  • Evaluate AI-generated scientific responses for accuracy and reasoning in biology.
  • Fact-check technical claims from public databases like PubMed and NCBI.
  • Assess experimental logic and annotate errors in biological sequences or protocols.

Prolific builds the largest pool of quality human data for AI development. With over 35,000 AI developers and researchers using the platform, it connects experts to train and evaluate AI models through ethical, paid participation.

$100,000–$200,000/yr

  • Review and critically evaluate new AI benchmarks on a regular cadence.
  • Produce clear, publication-ready research reports for broad audiences.
  • Analyze benchmark datasets using coding tools while maintaining analytical oversight.

The company focuses on producing rigorous, public-facing evaluations of AI benchmarks. It is a remote-first organization with a collaborative culture that values critical thinking and independence.

$100,000–$200,000/yr
Global 6w PTO

  • Review and assess new AI benchmarks at least every two weeks, evaluating their methodology and implications.
  • Publish and maintain public-facing reports on benchmarks, updating them as versions and models evolve.
  • Examine individual tasks within benchmarks in detail, using coding agents while maintaining critical oversight.

Epoch AI is a research institute investigating trends in machine learning and the economic consequences of AI. We aim to build a comprehensive knowledge base on AI, serving policymakers and society, with a small, focused team that values rigor and inclusivity.