Source Job

Global

  • Contribute to a cutting-edge AI benchmarking project by creating and reviewing high-quality, real-world natural sciences scenarios in Korean.
  • Adapt and apply clear scoring rubrics to evaluate AI-generated and human responses for accuracy and regulatory alignment.
  • Provide expert feedback and contribute to high-quality gold standard solutions for AI evaluation.

Content Evaluation Attention To Detail

10 jobs similar to Subject Matter Expert – Natural Sciences (Korean)

Jobs ranked by similarity.

Global

  • Design realistic scenarios in your target language or English grounded in operational contexts.
  • Adapt structured evaluation rubrics and review AI/human responses for accuracy, quality, and cultural appropriateness.
  • Contribute to gold-standard solutions reflecting best practices across target locale and domain.

LILT is an AI language company that provides multilingual AI and human-verified services to enterprises, governments, and AI developers. It operates with a global community of linguists and subject matter experts focused on innovation and excellence.

Global

  • Contribute to AI model training and evaluation in your area of expertise, including writing, reviewing, and assessing responses.
  • Evaluate AI-generated outputs for accuracy, logic, and nuance, and provide actionable feedback for model improvement.
  • Apply PhD-level judgment to real-world tasks in social sciences, humanities, arts, or linguistics.

Welo Data, part of Welocalize, is a global AI data company that delivers high-quality, ethical data to train advanced AI systems. With a community of over 500,000 contributors in 100+ countries, they emphasize flexibility, growth, and support for their contributors.

Global

  • Create and review realistic professional services scenarios in Nepali or English for AI benchmarking in Indian corporate contexts.
  • Adapt evaluation rubrics for analytical reasoning, technical problem-solving, and project coordination tasks.
  • Review AI and human-generated responses for factual accuracy, professional standards, and operational realism.

LILT provides multilingual AI and human-verified services to enterprises and governments worldwide. The company fosters a global, innovative community of linguists and subject matter experts dedicated to advancing human knowledge.

US

  • Perform side-by-side comparisons of AI-generated responses and evaluate them for factual accuracy, relevance, and overall quality.
  • Apply deep Korean expertise to assess language, terminology, tone, and cultural context specific to Korea.
  • Complete evaluations within established time and productivity expectations while maintaining consistent judgment.

Blueprint is a technology solutions firm that helps organizations turn complex challenges into meaningful outcomes across AI, cloud, data, and emerging technology. The company has teams across the United States and a culture built on high standards, ownership, and mutual support.

Global

  • Evaluate prompts and AI-generated outputs for accuracy, clarity, and cultural appropriateness.
  • Review and correct text, analyze multimedia content, and contribute voice recordings.
  • Apply careful judgment to ensure high-quality results aligned with task objectives.

LILT provides multilingual AI and human-verified services to enterprises, governments, and AI developers. The company has a global community of linguists and language professionals committed to innovation and excellence.

Global Europe

  • Contribute to shaping safer, smarter AI by joining a global network of linguists and culturally aware contributors.
  • Work on flexible, remote projects in annotation, evaluation, and prompt creation, always on your terms.
  • Get first access to projects that match your skills, from short tasks to multi-week assignments.

Welo Data, part of Welocalize, is a global AI data company with a network of over 500,000 contributors. They build smarter, more human AI by offering flexible, remote projects to a diverse community in over 100 countries, emphasizing growth and work-life balance.

Global

  • Design and build rigorous, verifiable Terminal-Bench tasks that test multilingual robustness in LLMs across prompt language effects and encoding edge cases.
  • Create realistic task environments with datasets and files in your native language, ensuring assets remain in the target language to genuinely measure multilingual handling.
  • Calibrate task difficulty by analyzing execution logs and participate in a 4-layer human quality control process to ensure benchmark integrity.

LILT is an AI and language technology company whose mission is to make the world's information available to everyone, regardless of language. They operate with a global community of linguists, engineers, and subject matter experts, fostering a culture of innovation and excellence.

Germany

  • Label, annotate, and evaluate German-language content including photos, graphics, and videos for linguistic and cultural accuracy.
  • Evaluate AI-generated content against Canva's quality bar for German users to shape language experiences.
  • Build and contribute to German-specific datasets to support the internationalization of Canva AI features.

Canva is a design platform redefining how the world experiences design. It is a global company with a large user base, known for its innovative culture and focus on AI-powered features.

Global

  • Design realistic healthcare and social assistance scenarios reflecting clinical and administrative settings in German-speaking locales.
  • Develop structured evaluation rubrics and review AI responses for medical correctness, operational feasibility, and patient safety.
  • Ensure cultural and contextual appropriateness of healthcare content, including hospital workflows and regulatory expectations.

LILT provides multilingual AI and human-verified services to Enterprises, Governments, and AI Developers worldwide. They have a global community of linguists and subject matter experts who collaborate on innovative projects advancing human knowledge.

Global

  • Create realistic, domain-specific mathematics tasks in Arabic reflecting local practices.
  • Adapt and apply clear scoring rubrics to evaluate AI-generated and human responses.
  • Review submissions and provide expert feedback for high-quality gold standard solutions.

LILT provides multilingual AI and human-verified services to Enterprises, Governments, and AI Developers worldwide. They have a global community of linguists and subject matter experts who thrive on innovation and excellence.