Senior AI/ML Test and Evaluation Engineer

OpenTeams

Remote regions

US

Salary range

$145,000–$250,000/yr

Benefits

Unlimited PTO

Similar Jobs

See all

Who We Are:

  • OpenTeams exists to make ownership possible by helping enterprises and governments build AI they control, govern, and evolve themselves.
  • Founded by Travis Oliphant, creator of NumPy and SciPy, our team has deep roots across the open-source ecosystem including NumPy, SciPy, PyTorch, and Jupyter.

About the Role:

  • You will build and operate the benchmarking and evaluation capability at the core of an AI platform.
  • You develop repeatable methodologies for comparing performance against current operational baselines and document limitations and failure modes.
  • Your reports inform senior stakeholder decisions on which capabilities are ready to field.

Key Responsibilities:

  • Design, implement, and operate benchmark execution and evaluation harnesses for AI models and agentic workflows.
  • Curate and recommend candidate benchmarks based on mission needs and document provenance of ground-truth and reference data.
  • Support partner organizations as they integrate capabilities with shared evaluation standards.

OpenTeams

OpenTeams helps enterprises and governments build AI they control, govern, and evolve themselves. Founded by NumPy and SciPy creator Travis Oliphant, the company is built by people with deep roots in the open-source ecosystem.

Apply for This Position