Source Job

$150,000–$250,000/yr
Global 5w PTO

  • Lead and accelerate the Applied White-Box Methods team's research agenda on AI safety.
  • Develop and evaluate white-box methods on realistic large-scale agentic coding tasks.
  • Publish findings and engage with the AI alignment community and external partners.

AI Safety Reinforcement Learning

20 jobs similar to Research Scientist

Jobs ranked by similarity.

$130,000–$200,000/yr
US

  • Develop and improve core AI methods and systems for reliable AI agents across the full lifecycle.
  • Create novel approaches for simulation, evaluation, and optimization of agent behavior in production.
  • Turn research ideas into working prototypes and production-facing capabilities.

This is an early-stage AI infrastructure company focused on making AI agents reliable in production. The company values innovation and practical deployment, with a small team driving frontier AI research and product development.

US

  • Craft creative prompts and multi-turn scenarios to stress-test AI guardrails across diverse risk categories.
  • Discover ways around safety filters and restrictions using jailbreak, evasion, and prompt injection techniques.
  • Evaluate and score model responses against structured harm taxonomies and severity rubrics.

Handshake AI partners with leading AI research labs to make models safer and more robust. Our red teaming operations help identify vulnerabilities before they reach users, contributing directly to the responsible development of frontier AI systems.

$105,000–$200,000/yr
Global

  • Support a portfolio of SPAR mentors across the round, helping them scope projects, set expectations, and keep teams on track.
  • Coach first-time mentors on running effective meetings, giving feedback, and managing mentees with different experience levels.
  • Handle sensitive team situations and advise on research output venues, while contributing to how SPAR recruits and supports mentors.

Kairos is a nonprofit accelerating talent into AI safety and policy. It's a small, high-trust team that has trained 1,200+ people in two years and expects to double in size.

Global

  • Build and refine models of AI takeoff and estimate their key parameters from public data and new experiments.
  • Work with engineers to execute large-scale experiments on frontier models and design proposals for safely pacing automated AI R&D.
  • Publish research papers and engage with academic, policy, and industry communities to help prepare for AI's transformative effects.

P-Zero Research is a public benefit corporation working to improve the long term trajectory of artificial intelligence. It aims to forecast and mitigate the risks of automated AI R&D and values alignment with its mission to keep AI safe.

$230,000–$330,000/yr
US Unlimited PTO

  • Own Merlin's foundation and world-model work, including architecture selection, post-training, and capability roadmap.
  • Lead and mentor a small team of world-model engineers, setting the technical bar and review culture.
  • Design model interface to the autonomy stack with structured, schema-constrained plan outputs and build evaluation harnesses.

Merlin is a publicly traded aerospace and defense company building a non-human pilot for full-stack aircraft autonomy. Headquartered in Boston, it is expanding its organization to accelerate the deployment of its autonomy platform.

UK US

  • Turning measurements of automated AI R&D into concrete policy proposals for decision makers.
  • Designing and running independent audits of frontier AI developers from scoping to defensible findings.
  • Working with high agency and comfort with ambiguity to advance AI safety.

P-Zero Research is a public benefit corporation improving the long-term trajectory of AI by forecasting and mitigating risks from automated AI R&D. The team values high agency, clear communication, and alignment with its mission.

$204,000–$290,000/yr
US Unlimited PTO

  • Lead Affirm's enterprise AI security review process, evaluating architecture, data flows, and design of AI tools and agentic systems.
  • Threat model AI/LLM systems for risks like prompt injection, insecure output handling, and data poisoning, and drive remediation.
  • Build security guardrails, tooling, and policy-as-code to automate AI security and support cross-functional initiatives.

Affirm is a financial technology company that offers clear, predictable point-of-sale installment loans with no hidden fees. The company is remote-first and values transparency, care, and flexibility, with a focus on building a diverse and inclusive team.

India

  • Build reinforcement learning and agent environments for real customer and Frontier lab use cases, including task specifications, scoring, and evaluation.
  • Develop benchmarks and evaluation harnesses to measure model and data quality across accuracy, robustness, safety, latency, and cost.
  • Run fine-tuning, adapter, and other model experiments to evaluate how data and methods influence model behavior, and deploy local or self-hosted models for evaluation and inference.

Appen has been a leader in AI training data for over 30 years, specializing in human-generated data to train, fine-tune, and evaluate models across generative AI, LLMs, computer vision, and speech recognition. They support model development through an AI-assisted data annotation platform and a global crowd of over 1 million contributors in more than 200 countries, fostering a culture of innovation, collaboration, and humility over ego.

$80,000–$110,000/yr
US

  • Run open-ended research projects using the internet as the primary tool.
  • Turn messy findings into clean outputs: matrices, briefs, landscape maps.
  • Use AI tools aggressively and build lightweight tools when needed.

A private design studio building and supporting a portfolio of companies across software, hardware, hospitality, and more. An affiliate of Expa, it also runs an early-stage investment fund and a founder cohort program.

$245,000–$307,000/yr
US

  • Lead the Sysdig Threat Research Team, owning adversary research and detection engineering end-to-end.
  • Drive AI security research into threats on models, agents, and infrastructure, plus AI-driven research methods.
  • Publish original research, ship detection content to customers and the Falco community, and partner with marketing and product teams.

Sysdig is a cloud security company that created Falco, the open standard for cloud threat detection, trusted by over 60% of the Fortune 500. With a culture recognized as a Best Place to Work and one of Deloitte's fastest-growing companies for 5 years, they emphasize diversity and open dialogue.

US 4w PTO

  • Build and ship AI methods for federal research data with LLMs, NLP, RAG, knowledge graphs.
  • Benchmark AI against baselines and human reviewers to prove gains in accuracy, reliability, efficiency, cost.
  • Set standards for trustworthy AI and mentor data scientists through code review and guidance.

ARI builds AI methods for federal research-funding and administrative data, with clients including the NIH. It is a small multidisciplinary team of scientists, analysts, economists, and data scientists experienced with NIH data.

North America Unlimited PTO

  • Study how engineers and customers use Archie in the field, identify where it succeeds or fails, and turn observations into actionable evaluations.
  • Recreate real-world engineering tasks and failure modes in repeatable environments for development teams.
  • Build and refine evaluation methodologies, including automated judges and human evaluation processes, to align with expert judgment.

P-1 AI is building Archie, an AI engineer agent for the physical world that works alongside human engineering teams. The company recently raised a $50 million Series A led by NEA and is driven by the mission of building superintelligence for engineering.

Global Unlimited PTO

  • Own White Circle's developer relations for how engineers and AI researchers discover and adopt our platform.
  • Lead technical launches, publish technical content, and run hackathons and meetups in SF and New York.
  • Build the technical community and partnerships across the AI ecosystem while tracking adoption metrics.

White Circle is an AI safety company building the safety, reliability, and optimization layer for AI systems with natural-language policies tested and enforced at scale. We are a small, focused team with $70M raised and 100M+ API calls monthly.

$130,000–$200,000/yr
US

  • Develop and improve AI methods, algorithms, and systems across the lifecycle of reliable AI agents.
  • Create new approaches for simulating, evaluating, and optimizing agent behavior in real-world settings.
  • Turn research ideas into working systems, prototypes, and production-facing capabilities.

The company builds infrastructure that helps enterprises make AI agents more reliable in production. It is a technical and research-focused team working on AI methods and practical systems.

US

  • Collaborate with client teams to diagnose operational bottlenecks and develop testable hypotheses.
  • Build and test AI-powered prototypes using coding agents like Claude Code or Cursor.
  • Drive adoption by working directly with users and iterating on solutions until they are effective in live workflows.

Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. The platform uses AI to review applications and provide a shortlist to employers, emphasizing efficiency and objectivity.

Unlimited PTO

  • Own the research roadmap for recursive self-improvement and ship what works into production.
  • Design agent strategies, exploration policies, and evaluations for autonomous LLM research systems.
  • Partner with applied scientists and build production-quality experiments end to end.

Sequen builds Recursive Ranking Intelligence (RRI), an autonomous engine where AI agents perform ML research in clients' clouds or on Sequen-hosted instances. It is a small, highly technical, early-stage team turning AI advances into production ranking systems.

$170,000–$170,000/yr
Global 7w PTO

  • Lead the engineering foundation behind reliable, measurable, and scalable AI-powered features.
  • Build evaluation frameworks, observability tooling, and diagnostic infrastructure for AI agents in production.
  • Combine hands-on software engineering with technical leadership and people management.

The company builds AI-powered products with a focus on reliability, scalability, and evaluation infrastructure. It operates as a fully remote, globally distributed team with an ownership-oriented culture.

India

  • Build RL environments, agentic systems, LLM pipelines, and evaluation frameworks for real-world AI use cases.
  • Develop benchmarks and evaluation harnesses to assess model quality across accuracy, safety, latency, and cost.
  • Conduct fine-tuning and model experiments, deploy self-hosted models, and document reproducible methodologies.

The employer is an organization focused on applied AI research, building practical and reusable AI systems for real-world use cases. It values curiosity, accountability, innovation, collaboration, and continuous learning in a remote environment.

$230,000–$322,000/yr
US

  • Lead the development and optimization of machine learning models to detect and prevent AI security risks like prompt injection and jailbreaks.
  • Build reproducible training and evaluation pipelines on Reddit's ML platform, partnering with platform engineers to improve performance and reliability.
  • Set the technical vision and multi-quarter modeling roadmap, mentoring engineers and establishing best practices for responsible ML development.

Reddit is a community of communities, built on shared interests and authentic conversations, with 100,000+ active communities and 130 million daily active visitors. It is one of the internet's largest sources of information, fostering a culture of openness and trust.

US 4w PTO

  • Own the technical direction of a product area, defining how teams specify, delegate, and verify agentic work.
  • Set standards for specification quality, verification, and security architecture to keep AI-generated work safe at scale.
  • Mentor engineers across teams and drive large-scale initiatives combining human and agentic contributors.

BambooHR builds a people intelligence platform that transforms HR for small and mid-sized businesses. The company is a values-driven market leader with a culture that champions growth, flexibility, and meaningful work.