Develop scalable ML pipelines across the full lifecycle and champion responsible AI for content understanding models and signals in production.
Provide technical leadership and mentorship to ML engineers and software engineers, setting technical standards and conducting design reviews.
Build evaluation and quality monitoring systems for content understanding signals using state-of-the-art LLM-as-judge practices.
Reddit is a community of communities built on shared interests, passion, and trust, home to the most open and authentic conversations on the internet. With 100,000+ active communities and approximately 126 million daily active unique visitors, it is one of the internet's largest sources of information.
Develop and improve NLP systems and language model-powered experiences.
Fine-tune and optimize language models for domain-specific use cases and build evaluation frameworks.
Deploy and maintain production-grade ML systems on GPU infrastructure with a focus on scalability and safety.
BetterHelp removes barriers to therapy and makes mental health care accessible globally. Founded in 2013, it is now the world's largest online therapy service with over 30,000 licensed therapists, and it invests deeply in employee well-being and professional development.
Lead the development and optimization of Large Language Models and Mixture of Experts models.
Collaborate with cross-functional teams to integrate ML models into our platform and conduct cutting-edge research in machine learning.
Mentor junior engineers and contribute to the team’s knowledge sharing and best practices.
webAI is an end-to-end private AI platform that enables enterprises and governments to bring AI to their data, powering specialized intelligence trained on their own knowledge. The company is a dynamic, fast-growing team fostering an exciting and growth-oriented work culture, committed to truth, ownership, tenacity, and humility.
Build models and data products that go from prototype to production, including generative models and subscriber-behavior predictions.
Dig into large, messy datasets to uncover trends and patterns, and contribute to the core Python data science library.
Build LLM-powered pipelines and agents, with comprehensive evals to validate model responses.
DevSavant is an operating partner for startups and growth-stage companies, helping them turn ambition into execution. With over 8 years in venture-backed ecosystems, they are trusted to accelerate delivery and scale teams efficiently.
Develop and refine features for deep learning models using large-scale customer and behavioral datasets.
Implement model architecture changes informed by recent academic research from venues like NeurIPS.
Optimize model training pipelines for efficiency and scalability while collaborating with client teams.
OpenTeams builds AI that empowers, offering energy-efficient and cost-effective models with a commitment to open source. The company values freedom, teamwork, accountability, and quality, and reinvests 3% of profits into the open-source community.
Train, evaluate, and iterate on ML models for customer feedback, including custom fine-tuning pipelines.
Build and maintain LLM-powered features like retrieval pipelines and insight agents.
Design and run robust evaluation frameworks to measure model performance.
Chattermill helps large brands like Uber, Amazon, and Wise put customers at the center using AI. They offer a flexible, trust-based culture with a choice-first environment.
Develop and operate production-ready AI and machine learning systems for enterprise-scale products.
Build and optimize LLM-powered applications, RAG pipelines, and intelligent agents.
Implement software engineering best practices for AI development including CI/CD and testing.
Our partner is building enterprise-grade AI solutions that deliver measurable business impact. They offer a remote-friendly work environment with a collaborative engineering culture focused on innovation, quality, and continuous learning.
Build and iterate on consumer-facing AI features powered by large language models (LLMs) and generative AI systems
Collaborate with engineers across the AI stack including prompt engineering and agentic workflow optimization
Run structured experiments and monitor production AI systems to optimize latency, cost, and scalability
Quora is a global knowledge sharing platform with over 300M monthly unique visitors, connecting people to share insights and learn. Poe provides a platform for users to chat and build with AI language models. They are a remote-first company with a culture rooted in transparency and experimentation.
Lead the design and operation of production machine learning systems for batch and online use cases with a focus on reliability and scalability.
Build and improve ML lifecycle infrastructure including training pipelines, inference workflows, monitoring, and automation.
Partner with cross-functional teams to translate business problems into ML solutions and guide prototypes to robust production systems.
Included Health is a healthcare company delivering integrated virtual care and navigation, aiming to raise the standard of healthcare for everyone. They are a remote-first organization offering comprehensive benefits and fostering a culture of inclusion.
Design and develop machine learning models for localization workflows, including machine translation and LLM finetuning.
Implement and optimize models using Python, TensorFlow, and deploy via Docker and AWS services.
Evaluate and select ML techniques, perform statistical analysis, and maintain clear documentation.
Welo Global is a leader in multilingual AI, technology, and content solutions serving over 2,000 clients in 300 languages. The company combines globally scaled multilingual infrastructure with a network of over 500,000 linguists and domain experts, backed by seven ISO certifications.
Design and deliver advanced AI solutions for legal and compliance operations at enterprise scale.
Build end-to-end AI applications using LLMs, RAG, and agentic frameworks with Python and modern orchestration tools.
Partner with legal stakeholders to identify automation opportunities and develop secure, scalable production systems.
Jobgether uses AI-powered matching to connect candidates with hiring companies, streamlining the application process. They operate remotely and focus on efficient, technology-driven recruitment.
Build and maintain scalable machine learning solutions in production.
Train and validate deep learning and statistical models for real-world applications.
Partner with product managers and engineers to define requirements and drive ML roadmap.
Twilio is a cloud communications platform that empowers businesses to build personalized customer experiences through APIs. With thousands of employees worldwide, the company champions a remote-first culture focused on inclusion and innovation.
Design and implement security controls for AI systems, including LLMs and ML pipelines.
Develop threat models and guardrails to protect against adversarial ML risks like prompt injection and model abuse.
Collaborate with cross-functional teams to ensure secure and responsible deployment of AI capabilities at scale.
Jobgether is a platform that uses AI-powered matching to connect candidates with job opportunities. They partner with companies to manage applications and next steps, focusing on efficient and fair hiring processes.
Design, build, and ship ML models that power content generation and quality eval scoring for Canva's generated element and template library.
Own the full ML lifecycle — from data pipelines and training through to deployment, monitoring, and iteration.
Partner with Content Engine, CORE AI Research, AI Media, and Discovery teams to align ML work with the broader content strategy.
Canva is redefining how the world experiences design with its intuitive design platform. We serve hundreds of millions of users globally and foster a culture of flexibility, inclusion, and innovation.
Lead independent test and evaluation research programs in NLP and LLM safety to protect the public from AI system harms.
Develop competitive research proposals, secure funding, and serve as principal investigator for digital safety research programs.
Analyze experimental results, design research challenges, and host workshops at leading AI safety conferences.
UL Research Institutes is a nonprofit organization dedicated to advancing safety science research to create a more secure and sustainable world. With a global reach and a century of experience, they foster a collaborative and scientific culture to address critical safety issues.
Own end-to-end Machine Learning (ML) system execution including data pipelines, training, and deployment.
Fine-tune and adapt models using state-of-the-art methods like LoRA and DPO.
Architect scalable inference systems and collaborate closely with application engineering.
This company develops advanced production-grade machine learning systems. The team is small and high-trust, with a culture of ownership and pragmatism.
Own end-to-end ML system execution including data pipelines, training workflows, evaluation systems, inference architecture, and deployment.
Fine-tune and adapt models using state-of-the-art methods such as LoRA, QLoRA, SFT, DPO, and distillation.
Architect scalable inference systems, balance latency, cost, and reliability, and deploy production-grade ML solutions.
Gina's Tech Jobs is a recruiting and staffing company that helps firms hire technical talent. They are a small agency focused on IT roles, fostering a high-trust, collaborative environment.
Develop and integrate ML, NLP, and Generative AI models to create smart collaborative solutions.
Build and optimize agentic workflows using LangChain, LangGraph, and similar frameworks.
Collaborate with engineering and product teams to deliver AI-driven features and prototype tools.
We are a technology company building next-generation AI-powered workplace assistants. Our team values innovation, engineering excellence, and continuous learning in an open and supportive environment.
Design and maintain LLM-powered backend services using Python and FastAPI.
Implement retrieval-augmented generation (RAG) for structured and unstructured fleet data.
Optimize retrieval accuracy, latency, and hallucination rates through automated evaluation pipelines.
Datakrew revolutionizes EV fleet intelligence with IoT and AI solutions. They aim to serve one million EVs within 5 years and cultivate a mission-driven culture.