Source Job

United States

  • Design, build, and maintain enterprise AI platform capabilities for LLMs, AI agents, RAG, and GenAI applications.
  • Develop reusable AI harnesses and evaluation frameworks to automate testing, model benchmarking, and quality assurance.
  • Implement observability, LLMOps pipelines, and scalable backend services using cloud-native and containerized architectures.

Python AWS Kubernetes Docker CI/CD

20 jobs similar to AI Platform and Harness Engineer

Jobs ranked by similarity.

$180,000–$210,000/yr
US

  • Contribute to the design and evolution of Expression's Agentic AI platform, defining scalable architectures for enterprise AI applications and services.
  • Architect, deliver, and optimize production-grade LLM services, agent workflows, orchestration layers, and retrieval pipelines.
  • Provide technical leadership and mentorship, guide architecture reviews, and raise the bar for engineering quality across the team.

Expression provides data fusion, analytics, software engineering, and spectrum management solutions to the US DoD and national security community. Founded in 1997 and headquartered in Washington DC, they have a "Perpetual Innovation" culture and were ranked #1 on Washington Technology's Fast 50 in 2018.

US

  • Design, develop, and deploy autonomous and multi-agent AI systems for reasoning, planning, and human-in-the-loop collaboration.
  • Engineer enterprise RAG pipelines with embeddings, hybrid retrieval, and prompt orchestration for explainable responses.
  • Build scalable backend services and APIs for enterprise AI workloads with security, reliability, and observability.

LTS builds an AI-native engineering platform to modernize legacy software systems, particularly for healthcare serving millions of Veterans. The engineering team is intentionally small, with every engineer having meaningful ownership and significant technical influence.

United States

  • Design and implement production-ready AI capabilities that improve reasoning, accuracy, and explainability.
  • Optimize autonomous agent workflows and Retrieval-Augmented Generation pipelines for performance.
  • Develop evaluation frameworks and benchmark datasets to measure AI effectiveness and drive continuous improvement.

LTS builds an AI-native engineering platform for modernizing mission-critical healthcare systems serving millions of Veterans. The engineering team is intentionally small, giving every engineer meaningful ownership and direct influence over product direction.

India

  • Design and implement scalable agentic solutions for diverse use cases.
  • Build and maintain a centralized AI platform with unified APIs for cross-functional team access.
  • Lead technical discussions, mentor junior engineers, and set the technical vision for the AI platform roadmap.

Cotiviti is a healthcare data analytics company that provides payment accuracy and quality improvement solutions. The company has a large global workforce and emphasizes a collaborative, innovative culture.

$153,000–$180,000/yr
US

  • Lead end-to-end development of advanced AI solutions including agentic systems, RAG pipelines, and multi-agent workflows.
  • Architect AI infrastructure with guardrails, access controls, and compliance protocols for ethical deployment.
  • Guide multi-disciplinary teams, setting technical playbook and best practices.

Nava is a consultancy and public benefit corporation that makes government services simple and effective. Since 2015, federal, state, and local agencies have trusted Nava to solve technology modernization challenges, and they maintain a collaborative, remote-friendly team environment.

US

  • Define and execute AI architecture strategies and technical roadmaps aligned with organizational goals.
  • Design scalable AI platforms leveraging Generative AI, LLMs, RAG, and automation.
  • Lead technical evaluations of emerging AI technologies and guide teams from proof-of-concept to production.

LTS is a technology solutions provider focused on AI and digital transformation for federal and enterprise clients. They foster a culture of innovation, growth, and collaboration.

US

  • Evolving the AI knowledge platform with retrieval, indexing, and synthesis for organization-wide use.
  • Architecting and operating agentic infrastructure on AWS with cost guardrails and observability.
  • Partnering with product engineering to define the AI platform API surface and building reference agent implementations.

ShiftKey is a healthcare workforce marketplace that connects facilities with licensed professionals to fill shifts, addressing staffing shortages. The company fosters an inclusive and collaborative culture, valuing diverse perspectives.

United States

  • Design and implement solutions in Python using AWS cloud computing capabilities, including developing AI and LLM-based workflows.
  • Operate in a collaborative, agile environment with a focus on enabling team success and creating proofs-of-concept and production solutions.
  • Engage with peers to ensure maintainable solutions and share knowledge across the company.

Two Six Technologies builds, deploys, and implements innovative products that solve complex challenges for US government and Fortune 50 clients. The company fosters a culture of collaboration and trust, empowering its team to push boundaries and support customers in building a safer global future.

US

  • Design and build agentic AI systems to automate complex workflows across ICON's software platform.
  • Develop LLM-powered features using state-of-the-art foundation models and APIs.
  • Architect multi-agent pipelines, tool-use systems, and MCP integrations.

ICON is a construction technology company that develops advanced 3D printing and robotic systems for building homes and structures. The company fosters an innovative, inclusive, and diverse work environment with a focus on pushing the boundaries of construction technology.

$200,000–$220,000/yr
US Unlimited PTO

  • Set technical direction across multiple services and teams as a senior individual contributor.
  • Define AI platform patterns including agent orchestration, tooling, governance hooks, and evaluation surfaces.
  • Stay hands-on, designing and building production systems in Python across the AI stack.

Mitratech builds world-class products that simplify operations in Legal, Risk, Compliance, and HR functions. Serving 20,000 client companies globally, including 30% of the Fortune 500, the company fosters a diverse, inclusive culture that blends entrepreneurial spirit with enterprise investment.

Global Unlimited PTO

  • Design and build scalable backend systems powering AI Agents in real-time enterprise environments.
  • Develop agent orchestration frameworks including multi-step reasoning, tool usage, and decisioning workflows.
  • Architect low-latency inference pipelines integrating LLMs, SLMs, and external tools at scale.

Level AI is an AI-native platform that transforms enterprise contact centers into engines of customer intelligence and business growth. Headquartered in Mountain View, CA, as a Series C company backed by Battery Ventures and ENIAC, their team builds agentic AI at scale.

US

  • Design and implement enterprise AI solutions using LLMs and RAG, focusing on healthcare applications for the VA.
  • Develop secure, scalable cloud-native applications with RESTful APIs on AWS, integrating with existing systems.
  • Collaborate with cross-functional teams to ensure AI accuracy, security, and compliance with Federal governance.

VetsEZ specializes in developing enterprise AI solutions for the Department of Veterans Affairs, focusing on secure, scalable healthcare applications. The company culture emphasizes collaboration with clinical stakeholders and adherence to Federal security standards, fostering a team of skilled engineers.

US Unlimited PTO 14w maternity 14w paternity

  • Work directly with customers to understand their workflows and operational constraints.
  • Build and deploy AI-enabled applications, agents, and workflow automation using LLM techniques.
  • Own delivery from prototype to production, including architecture, backend services, and integrations.

Trase Systems provides an end-to-end platform for deploying and optimizing AI in the enterprise, focusing on bridging the 'last mile' of AI adoption. As a startup founded in 2023, the team is small, agile, and collaborative, with a strong emphasis on innovation and direct customer impact.

Global

  • Design, build, and deploy production ML and LLM-based systems (RAG, agentic workflows, fine-tuning, embeddings) for enterprise clients.
  • Own technical delivery end-to-end: from architecture and prototyping to deployment, monitoring, and iteration.
  • Mentor and support other ML engineers on the team with code reviews, technical guidance, and knowledge sharing.

TensorOps is a boutique AI consultancy that bridges strategy and execution, designing and shipping production-grade AI systems for enterprise clients. We are a 100% remote team of 11+ people, partnering with unicorns and NASDAQ-listed companies, and have a culture of autonomy, open communication, and continuous learning.

$159,300–$230,000/yr
US

  • Architect and build end-to-end GenAI applications using Python, LangChain, and LlamaIndex on Google Cloud.
  • Develop advanced RAG pipelines and Semantic Search systems for production-level accuracy.
  • Optimize LLM and Embedding fine-tuning while applying MLOps best practices for scalability.

Egen is a fast-growing and entrepreneurial company with a data-first mindset, using advanced technology platforms like Google Cloud and Salesforce to drive client impact through data and insights. We are dedicated to learning and innovation, with a culture that values engineering expertise and solving tough problems.

US Unlimited PTO

  • Architect and own the enterprise AI data platform, including data ingestion, transformation, storage, and serving for all AI systems.
  • Design retrieval infrastructure for RAG applications, including embedding pipelines, vector stores, and hybrid search layers.
  • Own observability and evaluation frameworks for AI agents, ensuring output accuracy and audit trails.

3Pillar is an AI transformation partner helping enterprises build AI-native products and intelligent agents. With teams across multiple continents, they work with ambitious companies in financial services, healthcare, media, and technology.

US

  • Design, build, and maintain production-grade Generative AI, Agentic AI, and machine learning applications.
  • Develop evaluation and testing approaches to measure model performance and improve solution quality.
  • Partner with Engineering, Product, Data, and Operations teams to integrate AI capabilities into client’s systems.

We are an end-to-end AI transformation partner guiding enterprises from complex challenges to clear outcomes. We are a convergence of successful firms dedicated to fostering a culture of innovation and professional growth.

$90,000–$150,000/yr
United States

  • Design and deliver production AI and agentic systems across document intelligence, workflow automation, and copilots.
  • Own architecture decisions for LLM-based systems, including retrieval, tool use, orchestration, memory, and evaluation.
  • Manage evals and observability for production AI, ensuring system accuracy and detecting regressions.

Maxwell is a mortgage technology and fulfillment company on a mission to make lending simpler, faster, and more accessible. It is a remote-first team that takes craft seriously and moves with intention, building a cutting-edge AI company in mortgage technology.

US

  • Lead the design and delivery of complex enterprise-grade generative and agentic AI systems on AWS using services like Amazon Bedrock and SageMaker.
  • Architect scalable, secure, and production-ready AI platforms aligned with the AWS Well-Architected Framework.
  • Mentor engineers and collaborate with cross-functional teams to drive AWS AI practice evolution.

Robots & Pencils builds smart systems for a human world, blending creativity, engineering, and AWS-powered AI to help organizations reimagine how they work. They are a team of engineers and creatives with a culture of innovation and craftsmanship, focused on delivering enterprise-scale solutions.

Global

  • Design and build backend services for AI-powered product features, including inference pipelines and orchestration layers around LLMs.
  • Develop high-throughput, low-latency distributed systems with monitoring, logging, and alerting across production services.
  • Collaborate with product, infrastructure, and AI engineers to optimize performance, caching, batching, and streaming.

This team is building an AI-native productivity platform that replaces repetitive digital work with reliable AI workflows. They are a small, focused product team working on cutting-edge AI infrastructure.