Source Job

$45–$55/hr
US

  • Evaluate generated images against prompts for adherence, composition, realism, and technical defects.
  • Compare images side by side and select the stronger one with clear, evidence-based rationale.
  • Classify images against customer content policies covering sexual content, violence, and real-person likeness.

Generative AI Tools Content Moderation

9 jobs similar to AI Image Evaluator

Jobs ranked by similarity.

$45–$55/hr
US

  • Evaluate user requests and AI model responses against detailed customer policies with precise reasoning.
  • Distinguish subtle differences in context and intent to classify ambiguous cases accurately.
  • Participate in calibration discussions and contribute to improving evaluation frameworks.

Handshake powers a platform connecting 25 million job seekers with employers and educational institutions. Through Handshake AI, they provide data to frontier AI labs, having grown to a ~$1B run rate and paying over 30K individuals monthly.

$45–$55/hr
US

  • Evaluate user requests and AI model responses involving violence, weapons, threats, and dark fiction within the full conversation context.
  • Distinguish fictional, educational, historical, and defensive violence from requests that seek real-world uplift or express real intent to harm.
  • Assess whether a model's response gives meaningful real-world capability, and write concise rationales citing policy language and conversation details.

Handshake is a career platform founded on the belief that everyone deserves a path to a great career. Today, it powers 25 million job seekers, 1 million+ employers, and 1,600 educational institutions, with a rapidly growing AI data business that has reached a ~$1B run rate.

Global

  • Collect and curate fashion capsules according to established guidelines.
  • Research trends and build moodboards using references and AI generation tools.
  • Style models for photoshoots based on creative briefs independently.

BetterMe is an all-in-one well-being ecosystem that empowers millions worldwide to improve physically, mentally, and emotionally. With a culture of trust and transparency, 90% of leads are promoted internally, reflecting strong growth opportunities.

US

  • Support participants in virtual HR workshops by diagnosing and improving generative AI outputs.
  • Provide tool-agnostic guidance to help attendees refine prompts and achieve practical results.
  • Ensure responsible AI use and escalate issues as needed while fostering productive learning.

This company is a leading technology firm specializing in internet-related services and products, including search, cloud computing, and AI. It is a large global organization with a culture of innovation and collaboration.

US

  • Create short, narrated screen-capture walkthroughs demonstrating AI-tool tasks for college-level courses.
  • Use generative AI tools like ChatGPT, Claude, Gemini, and Copilot in practical professional workflows.
  • Structure video lessons with clear setup, logical sequencing, and a summary reinforcing key takeaways.

Study.com is the leading educational website providing lessons, courses, and practice for students, teachers, and adult learners. They empower millions of learners monthly and focus on making education accessible and increasing upward mobility through information.

India

  • Evaluate AI-generated content against domain-specific quality rubrics in humanities, arts, and culture.
  • Review documents, spreadsheets, and presentations for accuracy, relevance, clarity, and overall quality.
  • Provide structured feedback and collaborate with AI research teams to improve model outputs.

A partner company is seeking subject-matter experts to evaluate AI-generated content across humanities, arts, and culture. The company offers a flexible, remote contract environment, with no details on team size provided.

Canada

  • Evaluate AI-generated documents and presentations against quality standards.
  • Apply humanities expertise to identify inaccuracies and cultural issues.
  • Provide structured feedback to improve AI model performance.

A partner company is seeking a humanities evaluator to assess AI-generated content for accuracy and quality. The company emphasizes cultural awareness and critical thinking in a remote, asynchronous work environment.

Global

  • Review real user interaction traces with an AI shopping assistant
  • Identify logical failures, inaccuracies, or poor recommendations in the text
  • Create structured rubrics and verifiers to judge response quality

Terac is building the world's largest pool of vetted human experts for AI. Researchers, AI labs, and product teams use Terac to recruit, screen, and pay study participants across industries, languages, and skill sets.

Global

  • Audit multiple-choice options and correct answers for technical accuracy, eliminating ambiguous distractors.
  • Verify coding question prompts and grading rubrics, and write additional edge test cases.
  • Format and return the final corrected exam in a valid JSON structure.

Terac is building the world's largest pool of vetted human experts for AI. Researchers, AI labs, and product teams use Terac to recruit, screen, and pay study participants across industries, languages, and skill sets.