Source Job

20 jobs similar to Research Engineer / Research Scientist, Zero-Knowledge Verification

Jobs ranked by similarity.

UK US

  • Turning measurements of automated AI R&D into concrete policy proposals for decision makers.
  • Designing and running independent audits of frontier AI developers from scoping to defensible findings.
  • Working with high agency and comfort with ambiguity to advance AI safety.

P-Zero Research is a public benefit corporation improving the long-term trajectory of AI by forecasting and mitigating risks from automated AI R&D. The team values high agency, clear communication, and alignment with its mission.

EMEA

  • Build and maintain production-grade machine learning infrastructure for AI governance and trust, including model serving, evaluation, and monitoring.
  • Develop low-latency inference systems, model registries, and evaluation pipelines for reliable model lifecycle management.
  • Design and implement red-team testing, drift monitoring, and continuous retraining loops using telemetry and audit data.

The company builds AI governance and trust infrastructure, including production ML systems for AI agents. It offers a remote-first, international engineering organization with a focus on security and reliability.

$230,000–$322,000/yr
US

  • Lead the development and optimization of machine learning models to detect and prevent AI security risks like prompt injection and jailbreaks.
  • Build reproducible training and evaluation pipelines on Reddit's ML platform, partnering with platform engineers to improve performance and reliability.
  • Set the technical vision and multi-quarter modeling roadmap, mentoring engineers and establishing best practices for responsible ML development.

Reddit is a community of communities, built on shared interests and authentic conversations, with 100,000+ active communities and 130 million daily active visitors. It is one of the internet's largest sources of information, fostering a culture of openness and trust.

UK

  • Build tooling for capturing and processing data from agents and humans at significant scale.
  • Solve hard problems around compute, orchestration, scaling, security, and reliability.
  • Help develop approaches for training, benchmarking, and evaluating AI agents.

Prolific builds human data infrastructure for AI development, connecting researchers with a global pool of participants to collect high-quality, ethically sourced behavioral data. They are a mission-driven company at the forefront of AI innovation, with a remote culture and a focus on impactful work.

Spain

  • Design and develop large-scale platforms for LLM training and AI workloads.
  • Tackle distributed-systems challenges including intelligent job scheduling and resource optimization.
  • Collaborate with international teams to build production-ready AI infrastructure.

This role is with a partner company, an AI-focused R&D team building infrastructure for large language models. They are a fast-moving, highly technical team with a collaborative and innovative culture.

$176,100–$308,200/yr
North America Canada

  • Own a major subsystem of a novel exploitability engine end-to-end, including design, delivery, and quality.
  • Drive design and code reviews, raise the engineering bar, and mentor engineers through influence.
  • Partner with product, security R&D, and SecOps to turn customer problems into subsystem design.

ServiceNow provides an AI platform that automates workflows to free people from busywork, serving as the AI control tower for business reinvention. It serves 85% of the Fortune 500 and fosters an AI-native culture where technology and talent are unstoppable.

US Unlimited PTO

  • Build the foundations of the EdgeRunner Research organization, including data pipelines, model evaluation, and efficient parallelization.
  • Own codebases in areas like distributed training, quantization, compression, or compute cluster management.
  • Collaborate with a team of self-starters who operate independently and translate business objectives into technical solutions.

EdgeRunner AI builds state-of-the-art AI models for the tactical edge, enabling warfighters to make faster decisions and interact with robotic platforms. As a Series-A startup, we move quickly, fostering a culture of initiative, ownership, and comfort with ambiguity.

$180,000–$250,000/yr
Global

  • Work directly with leading AI labs and enterprises to define research goals and technical requirements.
  • Build data intelligence systems and implement ML pipelines for data curation, model training, and evaluation.
  • Develop LLM applications, including multi-agent systems, RAG workflows, and evaluation harnesses.

Our client is a venture-backed AI company building intelligent systems by combining human expertise with machine learning. With over $40 million in funding and a global expert network, they provide critical infrastructure for AI development.

Brazil

  • Design and develop specialized AI agents using Generative AI, LLMs, and agentic frameworks like LangChain and Semantic Kernel.
  • Build MCP servers and integrate AI applications with event-driven architectures, ensuring secure and traceable decisions.
  • Implement Human-in-the-Loop workflows and apply PromptOps practices for scalable, secure, and high-impact AI solutions.

The hiring company is a technology organization in Brazil focused on building AI-powered applications and agentic systems. It fosters a collaborative and international culture, with fully remote work and opportunities to work on innovative AI technologies.

Poland

  • Own the technical thesis for validating agent-generated software and build early proof of concept.
  • Design mid-loop feedback mechanisms that help agents self-correct toward working software.
  • Collaborate with engineering teams to adopt and measure trust in agentic development.

GitLab is an intelligent orchestration platform for DevSecOps that enables organizations to increase developer productivity, improve operational efficiency, and reduce security and compliance risk. With over 50 million registered users and trust from more than 50% of the Fortune 100, GitLab fosters a high-performance culture driven by values and continuous knowledge exchange.

Global

  • Build LLM-based agents on the platform's scaffolding, integrating tool calls, internal APIs, and guardrails.
  • Take agents to production on AWS with containers, CI/CD, secrets, permissions, and security controls.
  • Define and run evals, monitor with Langfuse, and document runbooks for independent operation.

Muttdata builds innovative Data Products and Machine Learning solutions to help companies solve complex business challenges. It is a fast-growing, remote-first startup that values collaboration, continuous learning, and a positive, ownership-driven culture.

APAC

  • Participate in code reviews of ERC-20, ERC-721, and other token smart contracts to identify security risks.
  • Explore and apply AI/LLM and Agent technologies to automate smart contract vulnerability detection.
  • Combine manual analysis with AI to analyze Web3 attack causes and impacts for audit process optimization.

Bybit is a leading cryptocurrency exchange and digital financial platform founded in 2018, serving over 80 million users across 200+ countries. Backed by a global team of ambitious builders and innovators, we foster a high-performance environment where talent drives real impact.

Global

  • Build and own AI backend and Agent workflows, including multi-turn conversations, RAG, tool calling, and handoff.
  • Drive evaluation and continuous improvement with product, design, and QA.
  • Own production reliability and diagnose issues across the stack.

Trust Wallet is the leading non-custodial crypto wallet, trusted by over 200 million people to securely manage digital assets. It has a global, fully remote team with a flat structure and a culture focused on learning and growth.

$204,000–$290,000/yr
US Unlimited PTO

  • Lead Affirm's enterprise AI security review process, evaluating architecture, data flows, and design of AI tools and agentic systems.
  • Threat model AI/LLM systems for risks like prompt injection, insecure output handling, and data poisoning, and drive remediation.
  • Build security guardrails, tooling, and policy-as-code to automate AI security and support cross-functional initiatives.

Affirm is a financial technology company that offers clear, predictable point-of-sale installment loans with no hidden fees. The company is remote-first and values transparency, care, and flexibility, with a focus on building a diverse and inclusive team.

$65,000–$97,000/yr
UK Europe Unlimited PTO

  • Collaborate directly with customers as an embedded technical expert, designing and deploying agentic AI solutions.
  • Quickly understand new industries, data, and systems to identify high-value AI and agentic workflows.
  • Build proof-of-concepts and production solutions using MCPs, agent frameworks, and tool-using LLMs.

We are a fast-growing company that helps customers turn AI ambitions into real production outcomes. We are fully remote with a 32-hour workweek and a collaborative, trust-based culture.

$235,000–$275,000/yr
Global

  • Break AI and agentic systems and turn research into automated, repeatable attacks for NodeZero.
  • Design and execute prompt injection, defense evasion, and tool-use exploitation against AI infrastructure.
  • Build and extend LLM-powered applications and microservices, focusing on production safety and reliability.

Horizon3 is a cybersecurity company whose NodeZero platform delivers autonomous pentests to find exploitable attack vectors. It’s a fast-growing, fully remote team of engineers and former special ops operators with a culture of respect and ownership.

EMEA

  • Lead technical evaluations and proofs of concept for an enterprise AI coding platform in real customer environments.
  • Engage directly with CTOs, VPs of Engineering, and senior developers on AI agent architecture and integration.
  • Manage enterprise security and deployment reviews, troubleshoot integration issues, and feed customer insights into product.

The company is an enterprise AI coding platform provider helping engineering organizations accelerate development with AI agents and LLM infrastructure. It operates as a remote-first global team with a fast-paced, innovative culture focused on collaboration and rapid product iteration.

US

  • Craft creative prompts and multi-turn scenarios to stress-test AI guardrails across diverse risk categories.
  • Discover ways around safety filters and restrictions using jailbreak, evasion, and prompt injection techniques.
  • Evaluate and score model responses against structured harm taxonomies and severity rubrics.

Handshake AI partners with leading AI research labs to make models safer and more robust. Our red teaming operations help identify vulnerabilities before they reach users, contributing directly to the responsible development of frontier AI systems.

US

  • Serve as a trusted advisor designing Cloudflare One architectures for Zero Trust and WAN.
  • Guide demos, proofs of value, and technical close while partnering with sales.
  • Build expertise in AI risks like Shadow AI and create reusable technical assets.

Cloudflare runs one of the world's largest networks, powering millions of websites for everyone from bloggers to Fortune 500 companies. We're a large global team that values AI-native curiosity and a culture of iteration.

$130,000–$200,000/yr
US

  • Develop and improve core AI methods and systems for reliable AI agents across the full lifecycle.
  • Create novel approaches for simulation, evaluation, and optimization of agent behavior in production.
  • Turn research ideas into working prototypes and production-facing capabilities.

This is an early-stage AI infrastructure company focused on making AI agents reliable in production. The company values innovation and practical deployment, with a small team driving frontier AI research and product development.