Source Job

8 jobs similar to AWS Trainium/NKI (TRN) AI Task Auditor

Jobs ranked by similarity.

Global

  • Evaluate AWS Serverless and IaC tasks for technical accuracy and reliability.
  • Provide clear, actionable feedback on deployment pipeline errors and architecture issues.
  • Ensure tasks are realistic, reproducible, and supported by robust tests.

We source experienced technical specialists to audit tasks used to train AI systems. Our project ensures that AI training workflows are technically rigorous and accurate, with a focus on high-quality deliverables.

Global

  • Evaluate Kubernetes tasks for technical accuracy, realism, and reproducibility.
  • Provide clear feedback on orchestration issues, configuration bugs, or logic errors.
  • Utilize your deep Kubernetes expertise to audit complex technical scenarios.

Greenhouse is a hiring platform that powers recruitment for modern companies. They are an established firm with a distributed team and a culture focused on innovation and flexibility.

Global

  • Assess software engineering tasks for technical accuracy, realism, and reproducibility.
  • Provide actionable feedback on codebase integration issues and logic errors.
  • Ensure AI training workflows are rigorous and practically applicable.

Project World Wide sources experienced technical specialists for AI training task auditing. This freelance contract opportunity focuses on ensuring technical rigor and accuracy in AI workflows.

Canada

  • Evaluate AI-generated documents, spreadsheets, and presentation decks against quality rubrics.
  • Identify factual, formatting, visual, and structural issues in professional deliverables.
  • Provide clear, structured feedback to enhance AI output quality and consistency.

This partner company specializes in AI training and evaluation, focusing on improving the quality of AI-generated professional content. Operating as a remote and asynchronous team, they value precision, collaboration, and independent work.

US

  • Evaluate AI-generated coding interactions end to end for usefulness, accuracy, and consistency with strong engineering practices.
  • Assess whether coding agents demonstrate sound technical reasoning and practical engineering judgment rather than just producing working-looking code.
  • Provide actionable feedback, distinguishing between adequate and exceptional AI response quality to shape evaluation standards.

Jobgether uses an AI-powered matching process to ensure your application is reviewed quickly and fairly against the role's core requirements. They are a third-party recruitment platform that partners with companies to fill positions, with a streamlined selection process.

  • Write detailed outlines of your regular workflows, focusing on one critical task performed at least weekly.
  • Provide structured evaluation tasks and nuanced feedback to train AI models.
  • Complete paid tasks remotely on a freelance basis, with most tasks requiring one hour of uninterrupted work.

Prolific builds the world's largest pool of quality human data for AI development. With over 35,000 AI developers and organizations using its platform, it focuses on ethically sourced behavioral data from paid participants.

US

  • Review and advise on evaluation criteria and scoring rubrics for AI-generated outputs.
  • Create, edit, and validate high-quality benchmark tasks and reference data for AI training.
  • Analyze model failures, including hallucinations and flawed reasoning, providing expert explanations.

The company specializes in AI evaluation and training, helping define standards for next-generation AI systems. They operate as a partner company that manages applications and next steps, valuing autonomy and expertise in their consultants.

Global

  • Evaluate LLM responses for accuracy, clarity, and completeness.
  • Fact-check technical claims using authoritative references.
  • Validate code and outputs, and annotate model performance.

Prolific builds the largest pool of high-quality human data for AI development, serving over 35,000 AI developers, researchers, and organizations. They connect researchers with a global community to collect ethically sourced behavioral data.