Build and implement AI features by selecting the right model and approach for each use case.
Evaluate and monitor AI feature performance in production to ensure accuracy and reliability.
Improve and iterate on prompts and implementations based on how features behave in the wild.
Aphex is a construction execution platform that replaces traditional spreadsheets with collaborative tools for delivery teams. They are a remote-first company with a growing engineering team in the Philippines, serving major contractors on multi-billion dollar projects.
Design, build, and ship production-grade AI agent features (PM Agent, Execution Agent, External Onboarding Agent) that automate real implementation work.
Lead design partner engagements to develop custom agent workflows, iterating based on customer feedback to productize reusable components.
Contribute to the MCP server and platform layer, ensuring reliable retrieval, evals, and tool integrations across the full AI stack.
TaskRay is the leader in post-sale work management in the Salesforce ecosystem, helping companies ensure flawless customer experiences post-sale. The company has a culture centered on Connection, Integrity, Hunger, and Thrive, with a small team that values high performance and innovation.
Own end-to-end Machine Learning (ML) system execution including data pipelines, training, and deployment.
Fine-tune and adapt models using state-of-the-art methods like LoRA and DPO.
Architect scalable inference systems and collaborate closely with application engineering.
This company develops advanced production-grade machine learning systems. The team is small and high-trust, with a culture of ownership and pragmatism.
Design and deliver production AI and agentic systems across document intelligence, workflow automation, and copilots.
Own architecture decisions for LLM-based systems, including retrieval, tool use, orchestration, memory, and evaluation.
Manage evals and observability for production AI, ensuring system accuracy and detecting regressions.
Maxwell is a mortgage technology and fulfillment company on a mission to make lending simpler, faster, and more accessible. It is a remote-first team that takes craft seriously and moves with intention, building a cutting-edge AI company in mortgage technology.
Design, develop, and deploy production-grade AI-powered backend systems integrating LLMs and traditional ML models.
Integrate and optimize vector databases for RAG pipelines, and write clean, well-structured Python code.
Debug complex cross-layer issues and collaborate with product and engineering teams to deliver cohesive solutions.
We are a fast-growing product company integrating cutting-edge AI capabilities into our core offering to deliver exceptional value to customers. Our small, fast-moving team works on practical, real-world AI applications with high autonomy.
Design and implement AI capabilities for intelligent data characterization and decision support.
Evaluate, optimize, and deploy open-weight foundation models for resource-constrained edge environments.
Develop efficient inference pipelines and implement RAG, semantic search, and model optimization techniques.
Expression provides data fusion, analytics, AI/ML, software engineering, and spectrum management solutions to the U.S. Department of Defense and national security community. Founded in 1997 and headquartered in Washington DC, the company was ranked #1 on Washington Technology's 2018 Fast 50 and is a Top 20 Big Data Solutions Provider, fostering a collaborative culture with opportunities for growth.
Build and improve ML components across data, training, evaluation, and inference.
Implement evaluation and testing to understand model behavior.
Debug model issues, performance problems, and production incidents.
This company builds core ML components for large-scale production systems. They emphasize real-world learning, iteration, and collaboration with senior engineers.
Build AI-powered product features integrated into real insurance workflows.
Design and optimize LLM-based interactions for customer and internal systems.
Improve reliability of AI outputs through guardrails, fallback logic, and validation layers.
BJAK is Southeast Asia's largest digital insurance platform, using AI to simplify insurance and financial services for millions of users. The company values technical excellence, speed of execution, and practical decision-making, with a global engineering team that works closely across product, design, and AI teams.
Develop and operate production-ready AI and machine learning systems for enterprise-scale products.
Build and optimize LLM-powered applications, RAG pipelines, and intelligent agents.
Implement software engineering best practices for AI development including CI/CD and testing.
Our partner is building enterprise-grade AI solutions that deliver measurable business impact. They offer a remote-friendly work environment with a collaborative engineering culture focused on innovation, quality, and continuous learning.
Design, build, and deploy production AI applications, copilots, retrieval systems, and agentic workflows.
Develop backend services, APIs, and application architectures integrating AI capabilities into enterprise systems.
Deploy AI solutions with security, observability, monitoring, evaluation, and governance.
Aimpoint Digital is a data, AI, analytics, and operations research advisory and solution engineering firm that helps organizations design and deploy enterprise-grade AI platforms and production applications. The company is a technical partner focused on moving beyond experimentation, emphasizing strong engineering discipline and scalable architecture.
Build AI-powered workflows, assistants, agents, and automation systems across customer support, CRM, and finance operations.
Work with product and engineering teams to turn manual processes into scalable AI-native systems using LLMs, APIs, and internal data.
Prototype quickly, test with users, then productionize what works, ensuring evaluation, monitoring, and reliability.
Bjak believes people deserve smarter ways to plan, save and grow their money, starting with the first mobile-first insurance platform in Southeast Asia. They have teams worldwide with over 20 nationalities working from offices and remotely, seeking talented and driven people.
Design, build, and operate AI systems that solve real business problems across Finom, from prototype to production.
Own AI systems end-to-end including problem framing, architecture, implementation, evaluation, deployment, monitoring, and iteration.
Partner with solution managers, domain teams, and engineers to integrate AI into real workflows and deliver production-grade AI capabilities.
Finom is a European tech startup developing an all-in-one financial B2B solution integrating banking, accounting, and invoicing. With over $346 million in total funding, they are expanding across key EU markets and maintain a start-up culture focused on innovation and employee empowerment.
Design and implement multi-agent state machines using LangGraph and LangChain for autonomous decision-making.
Build production-grade RAG pipelines on AWS with dual-engine vector search and serverless ingestion.
Architect secure multi-tenant AI systems with RBAC, PII redaction, and HITL guardrails.
Novara provides safety and operational risk management software that empowers organizations to identify and resolve issues before they become incidents. As a Providence Equity portfolio company, we are a dedicated team fostering innovation and collaboration in a remote-first environment.
Build, ship, and own product features end-to-end using cutting-edge AI/ML techniques.
Apply classical ML and LLM-based approaches like RAG, prompt engineering, and fine-tuning to enhance the audit and risk platform.
Collaborate with cross-functional teams in an Agile environment to deliver scalable, production-quality code.
Optro is a leading audit, risk, ESG, and InfoSec platform trusted by over 50% of the Fortune 500. The company has been named one of the 500 fastest-growing tech companies in North America for seven consecutive years, fostering a culture of innovation and collaboration.
Design, build, and deploy production ML and LLM-based systems (RAG, agentic workflows, fine-tuning, embeddings) for enterprise clients.
Own technical delivery end-to-end: from architecture and prototyping to deployment, monitoring, and iteration.
Mentor and support other ML engineers on the team with code reviews, technical guidance, and knowledge sharing.
TensorOps is a boutique AI consultancy that bridges strategy and execution, designing and shipping production-grade AI systems for enterprise clients. We are a 100% remote team of 11+ people, partnering with unicorns and NASDAQ-listed companies, and have a culture of autonomy, open communication, and continuous learning.
Build and ship AI agents, APIs, and applications on Affirm's internal platform, owning the full lifecycle from architecture to production.
Turn messy business requirements from People Operations stakeholders into production systems, integrating with tools like Workday and Notion.
Design reliability infrastructure for multi-model LLM services, including structured output validation and quality controls.
Affirm is reinventing credit to make it more honest and friendly, giving consumers the flexibility to buy now and pay later without any hidden fees or compounding interest. The People Tech & Analytics team builds and owns the data, AI, and technology infrastructure for Affirm's People function, running like a product engineering group embedded in HR.
Design and build the agent execution harness, owning the orchestration layer that routes inputs, manages context, and handles multi-step agentic workflows at scale.
Ensure reliability, observability, and fault tolerance in production AI systems while optimizing for latency and cost.
Lead eval engineering and prompt infrastructure to maintain agent quality and drive data-informed decisions across model updates.
ServiceNow is an AI platform company that automates workflows and helps businesses reinvent themselves through its AI control tower. It serves 85% of the Fortune 500 and fosters an AI-native culture focused on innovation and talent.
Design, develop, and deploy production-grade AI-powered backend systems.
Integrate large language models and machine learning models into scalable architectures.
Optimize system performance and implement strong testing practices.
Our partner company is building advanced AI-powered systems to create meaningful customer value. The team operates in a high-autonomy, fast-moving environment focused on production-ready AI solutions.
Own end-to-end ML system execution including data pipelines, training workflows, evaluation systems, inference architecture, and deployment.
Fine-tune and adapt models using state-of-the-art methods such as LoRA, QLoRA, SFT, DPO, and distillation.
Architect scalable inference systems, balance latency, cost, and reliability, and deploy production-grade ML solutions.
Gina's Tech Jobs is a recruiting and staffing company that helps firms hire technical talent. They are a small agency focused on IT roles, fostering a high-trust, collaborative environment.