Source Job

Canada

  • Conduct hands-on adversarial testing across AI models, applications, and data pipelines to identify vulnerabilities.
  • Perform advanced red-team assessments including jailbreaks, guardrail bypass, and prompt injection analysis.
  • Produce detailed vulnerability reports and collaborate with AI safety and engineering teams to improve security.

Python PyTorch TensorFlow LLM Security

20 jobs similar to Adversarial Machine Learning Engineer - Red Teaming

Jobs ranked by similarity.

US

  • Deliver AI red teaming security assessments for enterprise clients, defining scopes and authoring detailed reports.
  • Build and improve frameworks and tooling to scale AI red teaming service delivery and automate repetitive tasks.
  • Develop novel red teaming methodologies for emerging AI modalities and stay ahead of the latest AI security threats.

Check Point protects over 100,000 organizations worldwide from cyber and AI-driven threats using a prevention-first approach. They are recognized by TIME, Newsweek, and Forbes for excellence and workplace culture.

Germany

  • Design, build, and operate AI security runtime controls across cloud platforms.
  • Develop and maintain security solutions for AI interactions, LLMs, and agent workflows.
  • Act as senior escalation point for complex AI security incidents.

Canada

  • Design, develop, and deploy production-grade AI-powered backend systems.
  • Integrate large language models and machine learning models into scalable architectures.
  • Optimize system performance and implement strong testing practices.

Our partner company is building advanced AI-powered systems to create meaningful customer value. The team operates in a high-autonomy, fast-moving environment focused on production-ready AI solutions.

$100,000–$150,000/yr
US

  • Design and implement security controls and governance frameworks for AI and machine learning systems.
  • Develop threat models and security architectures tailored to LLMs, AI applications, and training data pipelines.
  • Identify and mitigate AI-specific risks like prompt injection, model abuse, and data exfiltration.

Jobgether is a platform that uses AI-powered matching to connect candidates with job opportunities. They operate as a technology-focused company with a remote team, emphasizing fair and efficient hiring processes.

Global

  • Evaluate LLM architecture logic for technical accuracy and audit ML code and notebooks for efficiency.
  • Refine RLHF frameworks to align models with human intent and analyze model reasoning in complex chain-of-thought prompts.
  • Benchmark performance by conducting comparative testing between model outputs based on technical metrics.

Prolific connects researchers with a global pool of participants for collecting high-quality human data to train AI models. With over 35,000 users, they focus on ethical data gathering to advance AI capabilities.

Canada

  • Design and develop advanced AI agent platforms for enterprise automation and intelligent customer experiences.
  • Architect agent workflows incorporating reasoning, tool usage, retrieval, guardrails, and performance monitoring.
  • Build scalable Python services and AI platform components deployed in cloud environments.

Our partner is a technology company developing advanced agentic AI platforms for enterprise automation. They foster an inclusive, remote-friendly culture with a focus on innovation and collaboration.

Global

  • Design and build the AI execution platform with event-triggered workflows and model-agnostic runtimes.
  • Develop evaluation layers with golden test suites and safety checks to ensure AI reliability.
  • Implement governance mechanisms and build AI agents supporting business workflows.

Jobgether is a platform that connects talent with opportunities using AI-powered matching. The company has a globally distributed team and a remote-first culture.

US

  • Lead a high-performing AI and data science team to develop and deploy AI-driven security systems.
  • Define and execute the technical roadmap for AI, ML, and data systems, ensuring robust performance and scalability.
  • Manage the entire lifecycle of data pipelines and AI models, integrating with platforms like AWS Bedrock and OpenAI.

Bugcrowd is a crowdsourced security platform that unites organizations with a global network of hackers to identify vulnerabilities. The company is backed by prominent investors and fosters a collaborative, inclusive culture with a diverse team of professionals.

$216,700–$303,400/yr
US

  • Design, develop, and deploy ML models, including large language models, for various NLP tasks.
  • Collaborate with cross-functional teams to gather requirements, define architectures, and iterate on model development.
  • Stay up-to-date with latest research and contribute to best practices for responsible ML development.

Reddit is a community of communities built on shared interests, passion, and trust, hosting the most open and authentic conversations on the internet. With over 100,000 active communities and approximately 126 million daily active users, it is one of the internet's largest sources of information, fostering a culture of authenticity and community.

Global

  • Design, build, and deploy production ML and LLM-based systems (RAG, agentic workflows, fine-tuning, embeddings) for enterprise clients.
  • Own technical delivery end-to-end: from architecture and prototyping to deployment, monitoring, and iteration.
  • Mentor and support other ML engineers on the team with code reviews, technical guidance, and knowledge sharing.

TensorOps is a boutique AI consultancy that bridges strategy and execution, designing and shipping production-grade AI systems for enterprise clients. We are a 100% remote team of 11+ people, partnering with unicorns and NASDAQ-listed companies, and have a culture of autonomy, open communication, and continuous learning.

$190,000–$319,000/yr
US

  • Design and ship security workflows combining deterministic analysis with LLM reasoning to find real vulnerabilities across languages and frameworks.
  • Engineer agentic pipelines and prompts that are precise, cost-aware, and trustworthy for security-critical work.
  • Push on hard problems in automated triage and validation to close the gap between finding and actionable fix.

Semgrep is a code security platform that helps teams catch and fix vulnerabilities before they ship. They are a venture-backed startup with a transparent culture that values respect and honesty.

Canada

  • Own the roadmap for AI-powered development frameworks including agent workflows and evaluation systems.
  • Expand AI-enabled workflows across teams to improve efficiency and business outcomes.
  • Productize internal AI solutions by developing connector layers, skill libraries, and onboarding experiences.

A company focused on building AI-powered platforms to transform team productivity. The company operates remotely and values innovation and collaboration.

North America Unlimited PTO 12w maternity 8w paternity

  • Lead state-of-the-art research and development in AI/ML, advanced statistical modeling, and applied mathematics to solve security problems.
  • Develop novel AI workflows and models to address Ent's security challenges, incorporating and advancing current SOTA approaches.
  • Collaborate with the team to invent frameworks and theoretical grounding while thinking outside the box to combine multiple techniques.

Ent is an intent-aware workspace security platform that secures human and AI-driven work by understanding intent and intervening at the moment of risk. Founded by former RiskIQ co-founders, backed by leading VCs, and in production with Global 2000 customers, the company values a customer-first, humble, and urgent culture.

US

  • Design and build advanced AI-powered systems including autonomous agents and generative AI capabilities.
  • Lead architecture and implementation of intelligent systems that improve digital learning experiences.
  • Collaborate with cross-functional teams to bring AI solutions from concept to production.

The company builds AI-powered digital learning platforms that transform educational experiences. They serve millions of learners and educators with a focus on responsible AI and inclusive culture.

US

  • Fine-tune and optimize Large Language Models for healthcare and life science applications.
  • Develop and enhance RAG pipelines to improve AI-powered information retrieval.
  • Collaborate with cross-functional teams to deploy deep learning solutions for complex healthcare challenges.

The company develops AI-powered healthcare solutions using Large Language Models. They foster a highly collaborative and innovation-driven environment with a diverse international team.

US Unlimited PTO

  • Collaborate with product and engineering teams to translate product objectives into autonomous agent-based solutions.
  • Design and build new agent data types, pipelines, and frameworks to coordinate reasoning, function calling, and actions.
  • Develop and optimize autonomous agents leveraging LLMs, planning algorithms, and multi-step reasoning approaches.

PointClickCare is a leading health tech company that helps providers deliver exceptional care. As a founder-led, privately held company with over 30,000 provider organizations and 400+ integrated partners, they are recognized by Forbes as a top private cloud company and honored as one of Canada's Most Admired Corporate Cultures, offering flexibility and growth opportunities.

Brazil

  • Lead the design and development of scalable AI solutions, from experimentation to production deployment.
  • Define AI engineering standards and best practices, influencing architecture decisions across teams.
  • Collaborate with cross-functional stakeholders to integrate generative AI and LLM capabilities into products.

The company is at the forefront of AI-driven product development, focusing on building scalable and intelligent systems. It fosters a culture of innovation and technical excellence, with a remote team and a commitment to engineering leadership.

Europe

  • Define and execute the research strategy for core AI intelligence, including reasoning, planning, memory, and context representation.
  • Lead technical decisions on model architecture, balancing proprietary solutions with state-of-the-art models.
  • Establish and oversee AI alignment, safety, and guardrail strategies to ensure responsible system behavior.

This company develops an advanced AI platform to automate complex digital workflows. They are a small, highly experienced team with a fast-moving culture focused on cutting-edge machine learning.

US Unlimited PTO

  • Own the technical direction, architecture, and execution of HiddenLayer’s AI Attack Simulation product.
  • Partner with Product and Security Research to define product strategy and translate adversarial AI techniques into scalable capabilities.
  • Design and build agentic workflows, tool integrations, and evaluation frameworks while mentoring engineers.

HiddenLayer protects the world’s most valuable technologies from adversarial AI attacks. They are a venture-backed startup with a remote global team, recognized for innovation and committed to diversity and inclusion.

Global 4w PTO

  • Take ownership of protecting product, users, and collaborators as a Security Engineer on a fast-growing AI dating platform.
  • Perform weekly code reviews, coordinate security audits, and maintain risk register to ensure safety of AI-generated content.
  • Monitor emerging AI security threats, manage IT access and device security, and run phishing simulations company-wide.

EverAI builds the world's largest AI companionship platform, redefining relationships for millions with a proprietary moderation system, EverGuard. With 50 million users in two years, the team of about 100 is enthusiastic, passionate, and hardworking, led by experienced entrepreneurs.