Source Job

Canada

  • Evaluate AI-generated media, journalism, and communications content against quality rubrics.
  • Review outputs for factual accuracy, relevance, clarity, tone, and structure.
  • Provide structured, actionable feedback to improve AI model performance.

Journalism Communications Editorial Judgment Attention To Detail

20 jobs similar to Media / Journalism / Communications Evaluator

Jobs ranked by similarity.

Canada

  • Evaluate AI-generated documents and presentations against quality standards.
  • Apply humanities expertise to identify inaccuracies and cultural issues.
  • Provide structured feedback to improve AI model performance.

A partner company is seeking a humanities evaluator to assess AI-generated content for accuracy and quality. The company emphasizes cultural awareness and critical thinking in a remote, asynchronous work environment.

India

  • Evaluate AI-generated content against domain-specific quality rubrics in humanities, arts, and culture.
  • Review documents, spreadsheets, and presentations for accuracy, relevance, clarity, and overall quality.
  • Provide structured feedback and collaborate with AI research teams to improve model outputs.

A partner company is seeking subject-matter experts to evaluate AI-generated content across humanities, arts, and culture. The company offers a flexible, remote contract environment, with no details on team size provided.

Canada

  • Evaluate AI-generated documents, spreadsheets, and presentation decks against quality rubrics.
  • Identify factual, formatting, visual, and structural issues in professional deliverables.
  • Provide clear, structured feedback to enhance AI output quality and consistency.

This partner company specializes in AI training and evaluation, focusing on improving the quality of AI-generated professional content. Operating as a remote and asynchronous team, they value precision, collaboration, and independent work.

Canada

  • Evaluate AI-generated finance documents against quality standards and identify issues.
  • Review outputs for accuracy, consistency, and formatting errors.
  • Provide structured, actionable feedback to improve AI content.

The partner company develops AI systems and seeks evaluators to improve output quality. The work is remote and asynchronous, with a focus on independent contributions.

United States

  • Review search results and evaluate their relevance to user queries
  • Answer true/false questions about content quality
  • Rate search results based on guidelines to improve AI systems

Welo Data provides AI services and data validation to improve search engine and AI systems. They are a remote-first company with a focus on quality and support for their contractors.

US

  • Evaluate AI-generated content for quality, accuracy, and cultural relevance
  • Apply Castilian Spanish expertise to assess response appropriateness for Spain
  • Provide structured feedback and document decisions to improve AI performance

Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. They use objective, data-driven recruitment processes and prioritize privacy and fairness.

US

  • Review, evaluate, and annotate AI-generated content across text, images, audio, and video.
  • Perform quality checks to ensure accuracy, consistency, and compliance with project guidelines.
  • Identify edge cases and inconsistencies, contribute to high-quality dataset development, and participate in calibration activities.

Welo Data, part of Welocalize, is a global AI data company with over 500,000 contributors that provides high-quality, ethical data for training advanced AI systems. The company supports a diverse, global community across 100+ countries and offers project-based freelance opportunities with flexibility and growth potential.

US

  • Evaluate simulated advertiser-AI conversations for technical accuracy and campaign structure.
  • Fact-check platform strategies against correct hierarchy and full-funnel metrics.
  • Write clear, actionable feedback to correct errors and improve AI model performance.

RWS specializes in AI training data and language services, providing data annotation and evaluation solutions to improve AI model reliability. The company fosters a culture of diversity and inclusion, operating as a global employer with a focus on equal opportunity.

$4–$5/hr
Global

  • Review AI-generated responses against source images and quality guidelines.
  • Identify issues like hallucinations, missing details, or policy violations.
  • Provide structured feedback to improve model performance and output quality.

Jobgether uses AI-powered matching to connect candidates with partner companies. They focus on efficient, objective hiring processes and operate as a platform for remote opportunities.

Canada

  • Evaluate AI-generated legal and business documents against quality standards and apply professional judgment.
  • Review contracts, diligence materials, redlines for accuracy, consistency, and completeness.
  • Provide clear, structured feedback to improve AI-generated legal content.

$70,000–$85,000/yr
Global Unlimited PTO 13w maternity 13w paternity

  • Manage editorial workflows and coordinate communications between editorial teams and partners.
  • Assist with content reviews, editorial quality checks, and provide actionable feedback.
  • Maintain editorial resources, best practices, and support partner education initiatives.

Stacker is a platform that helps brands and nonprofit newsrooms extend the reach of their content by syndicating across a network of thousands of trusted news publishers. As a bootstrapped, fast-growing company, they are a remote-first team that values ownership, integrity, and collaboration, offering flexible schedules and unlimited vacation.

UK

  • Evaluate AI-generated French responses, rate them, and flag cultural issues.
  • Rewrite weak responses into clear, natural Canadian French.
  • Create original French prompts and example responses to expand training data.

We are a global AI data company that delivers high-quality, ethical data to train the world's most advanced AI systems. With over 500,000 contributors, we offer flexible, remote project-based opportunities with a supportive global community.

Canada

  • Create original, human-authored content in specialized domains like law, business, and philosophy for AI training.
  • Work independently from creative prompts to produce accurate, original written samples without AI tools.
  • Collaborate with a global community of writers and experts to improve language technology and AI evaluation.

Jobgether uses AI-powered matching to connect candidates with hiring companies, streamlining the application process. It operates as a platform that shares shortlisted candidates with employers, who manage final decisions and next steps.

$58,948–$65,883/yr
Canada

  • Review and edit all copy across email, paid social, digital display, web, and video scripts for five institutions, ensuring accuracy and brand voice consistency.
  • Partner with Legal to review content against regulations and compliance standards, flagging language risks and required disclosures.
  • Optimize content for SEO and AI search visibility, applying best practices for query alignment and generative engine optimization.

OLIVER is the world's first and only specialist in designing, building, and running bespoke in-house agencies and marketing ecosystems for brands. Partnering with over 300 clients in 40+ countries, we are part of The Brandtech Group and leverage cutting-edge AI technology to drive creativity and efficiency.

Global

  • Evaluate model-generated content across multiple modalities including text, images, audio, and video.
  • Apply defined quality rubrics such as factuality, consistency, and aesthetics to assess outputs.
  • Conduct independent research on unfamiliar topics to make well-supported evaluation judgments.

Our client is a global technology company that helps businesses build, train, and manage AI systems. They offer flexible, remote work with variable workload and a focus on high-quality model evaluation.

US Canada

  • Lead content strategy across jerry.ai and external platforms, owning organic traffic and conversion targets.
  • Design and run an AI-native content production pipeline with LLM prompts, QA gates, and human checkpoints.
  • Build authority in SEO, AEO (Answer Engine Optimization), and YMYL compliance for insurance and finance content.

Jerry.ai is building an AI agent to manage physical assets like cars and homes, starting with insurance and expanding to repairs and diagnostics. The company has reached 5M+ customers, raised $240M+, and became profitable in early 2024, with a remote-first culture and offices in Palo Alto, New York, Chicago, and Toronto.

$22–$22/hr
Canada

  • Evaluate and rank model outputs, stress-test models for failure modes, and create high-quality datasets with detailed rubrics.
  • Annotate and correct multimodal data, maintain consistency through calibration exercises, and adapt to evolving task types.
  • Report on model performance trends and provide clear feedback to cross-functional partners on model successes and failures.

Cohere is a security-first enterprise AI company that builds cutting-edge foundation models and end-to-end products for real-world business problems. It is a global technology company with offices in Toronto, San Francisco, London, New York, Montreal, Seoul, Germany, and Paris, staffed by a team of passionate researchers, engineers, and designers.

$80–$150/hr
UK

  • Review AI-generated responses to clinical scenarios for accuracy and safety.
  • Compare and justify the best responses among multiple model answers.
  • Write improved exemplars and structured feedback to enhance AI model learning.

Prolific is building the largest pool of quality human data in the world, used by over 35,000 AI developers and researchers. It is a platform that connects researchers with a global participant pool for ethically sourced human data.

Global

  • Evaluate prompts and AI-generated outputs for accuracy, clarity, and cultural appropriateness.
  • Review and correct text, analyze multimedia content, and contribute voice recordings.
  • Apply local insight into tone, symbolism, visual cues, and market fit to deliver culturally relevant content.

LILT is an AI company that makes the world's information available to everyone, no matter the language they speak. They work with a global community of linguists and subject matter experts to deliver multilingual AI and human-verified services to Enterprises, Governments, and AI Developers.

Global

  • Provide native-level Canadian French language vetting and QA for AI data projects.
  • Annotate and review AI outputs for grammatical accuracy, cultural context, and naturalness.
  • Develop educational resources and feedback documentation to improve AI alignment.

We are an AI training company that focuses on language alignment and data annotation for AI systems. Our remote team values linguistic precision and cultural nuance.