Remote Software engineering Jobs · PyTorch

Job listings

  • Lead the design and development of scalable AI solutions, from experimentation to production deployment.
  • Define AI engineering standards and best practices, influencing architecture decisions across teams.
  • Collaborate with cross-functional stakeholders to integrate generative AI and LLM capabilities into products.

The company is at the forefront of AI-driven product development, focusing on building scalable and intelligent systems. It fosters a culture of innovation and technical excellence, with a remote team and a commitment to engineering leadership.

  • Build and ship AI features end-to-end, from model to system to user experience.
  • Design and iterate on prompts, tools, memory, and agent workflows for real-world reliability.
  • Debug full-stack issues and optimize for latency, cost, and production performance.

A1 builds a proactive smart assistant for everyday users, bringing intelligence to conversations, errands, organizing, and workflows with minimal prompting. The team is small, world-class, and focuses on rapid iteration and shipping high-quality AI products.

  • Own the full post-training pipeline from data curation to deployment.
  • Advance techniques across the post-training stack including SFT, RLHF, DPO, and reward modeling.
  • Build personalization and customization capabilities for user adaptation.

Black Forest Labs is a research lab behind foundational generative AI technologies like Stable Diffusion and FLUX, powering tools used by millions worldwide. They are a fast-growing, distributed team with a culture of research excellence, open science, and low ego.

  • Design, build, and deploy production ML and LLM-based systems (RAG, agentic workflows, fine-tuning, embeddings) for enterprise clients.
  • Own technical delivery end-to-end: from architecture and prototyping to deployment, monitoring, and iteration.
  • Mentor and support other ML engineers on the team with code reviews, technical guidance, and knowledge sharing.

TensorOps is a boutique AI consultancy that bridges strategy and execution, designing and shipping production-grade AI systems for enterprise clients. We are a 100% remote team of 11+ people, partnering with unicorns and NASDAQ-listed companies, and have a culture of autonomy, open communication, and continuous learning.

$153,351–$206,481/yr

  • Lead a team to build, scale, and optimize the ML infrastructure powering drug discovery.
  • Collaborate with ML engineering, data science, and research teams to deliver scalable solutions.
  • Mentor and coach team members in MLOps, distributed computing, and infrastructure engineering.

Recursion is a clinical-stage TechBio company decoding biology to develop medicines. With a focus on AI and machine learning, the company fosters a culture of bold integrity and cross-functional collaboration.

$95,000–$195,000/yr

  • Develop and implement AI-based data assimilation systems to improve weather forecasting.
  • Design data pipelines and integrate diverse observational datasets for model training and evaluation.
  • Communicate findings and collaborate with NOAA scientists to ensure robust system deployment.

Lynker provides professional, scientific, and technical services, specializing in hydrology, geospatial analysis, IT, and resource management. As an employee-owned business, they foster a collaborative culture of skilled professionals focused on creative solutions.

  • Lead efforts in engineering, building, testing, and deploying advanced cloud-enabled AI/ML solutions for surveillance, targeting, and engagement operations.
  • Partner with project managers and engineering teams to define objectives for UI/UX systems and develop prototypes to address mission-specific requirements.
  • Conduct rigorous testing, analyze test data, and refine systems to ensure robustness, scalability, and cyber resilience.

Barbaricum is a rapidly growing government contractor providing leading-edge support to federal customers, with a particular focus on Defense and National Security mission sets. Founded in 2008, the company has over 17 years of experience and a corporate culture diverse in expertise and perspectives with a focus on collaboration and innovation.

  • Develop scalable ML pipelines across the full lifecycle and champion responsible AI for content understanding models and signals in production.
  • Provide technical leadership and mentorship to ML engineers and software engineers, setting technical standards and conducting design reviews.
  • Build evaluation and quality monitoring systems for content understanding signals using state-of-the-art LLM-as-judge practices.

Reddit is a community of communities built on shared interests, passion, and trust, home to the most open and authentic conversations on the internet. With 100,000+ active communities and approximately 126 million daily active unique visitors, it is one of the internet's largest sources of information.

US Unlimited PTO

  • Lead the development and optimization of Large Language Models and Mixture of Experts models.
  • Collaborate with cross-functional teams to integrate ML models into our platform and conduct cutting-edge research in machine learning.
  • Mentor junior engineers and contribute to the team’s knowledge sharing and best practices.

webAI is an end-to-end private AI platform that enables enterprises and governments to bring AI to their data, powering specialized intelligence trained on their own knowledge. The company is a dynamic, fast-growing team fostering an exciting and growth-oriented work culture, committed to truth, ownership, tenacity, and humility.

  • Architect and build large-scale ML systems spanning data, training, evaluation, inference, and deployment.
  • Implement evaluation pipelines covering performance, robustness, safety, and bias.
  • Own production deployment including GPU optimization, memory efficiency, latency reduction, and scaling policies.