Source Job

Global

  • Own the technical path from customer interest to working deployment, integrating the platform into production AI environments.
  • Build and operate AI/MLOps pipelines, debug complex environments, and create prototypes and demos.
  • Translate customer needs into product improvements, partnering with Sales, Product, and Engineering.

Python Kubernetes MLOps Cloud-Native

20 jobs similar to Founding Forward Deployed AI Engineer

Jobs ranked by similarity.

$170,000–$250,000/yr

  • Passionate about building AI models from the ground up and deploying them in production
  • Excited about joining early-stage startups and working directly with founders
  • Interested in shaping AI strategy and leading the development of AI-powered products

SignalFire partners with top early-stage startups shaping the future of technology. They have a portfolio of over 200 innovative companies across AI, cybersecurity, healthtech, fintech, developer tools, and enterprise SaaS.

North America

  • Build and deploy production code to support customer AI inference workloads on Tenstorrent's hardware and software stack.
  • Debug and optimize across the full inference stack, from serving layer to kernel dispatch, and translate customer issues into actionable requirements.
  • Operate Kubernetes and observability tools to manage multi-node AI clusters and ensure reliability.

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. Their diverse team of technologists has developed a high-performance RISC-V CPU from scratch, and they value collaboration, curiosity, and a commitment to solving hard problems.

$180,000–$200,000/yr
US

  • Design and build the company's AI-first platform architecture.
  • Develop agent-based systems, orchestration layers, and inference pipelines.
  • Build scalable data ingestion, transformation, and feature engineering pipelines.

Our client is an early-stage healthcare AI company leveraging AI, data science, and clinical expertise to generate actionable insights from health data. The small team works alongside experienced founders and world-class advisors, building an AI-native platform from scratch.

LATAM North America EMEA

  • Spend your first weeks in the operator's seat, learning the customer's job from the inside before writing any code.
  • Ship production GenAI/LLM systems that move business unit metrics, not just complete scope.
  • Work embedded in small, senior teams alongside Principal Architects, owning the outcome from start to finish.

Provectus is a Premier AWS partner and an Anthropic Strategic Partner at the forefront of applied AI, helping enterprises turn Claude, agentic systems, and their own data into measurable business outcomes. With offices in North America, LATAM, and EMEA, we partner with clients worldwide and our team holds 100+ AWS certifications and is Claude Code certified.

Europe

  • Monitor, operate, and support production AI infrastructure platforms including NVIDIA GPU environments.
  • Investigate and resolve infrastructure, networking, hardware, and platform-related incidents.
  • Collaborate with engineering teams, vendors, and datacenter personnel to improve operational processes.

Mirantis is the Kubernetes-native AI infrastructure company, enabling organizations to build scalable, secure, and sovereign infrastructure for modern AI and data-intensive applications. Serving enterprises like Adobe and PayPal, the company combines open source innovation with deep Kubernetes expertise to deliver composable developer platforms across any environment.

$125,000–$250,000/yr
Global

  • Design, operate, and improve reliable infrastructure for AI training and inference workloads.
  • Build monitoring, alerting, runbooks, and incident-response practices for easier operations.
  • Partner with ML, research, and platform teams to translate workload needs into infrastructure improvements.

Boson AI builds production-grade AI systems that make communication with AI more natural, capable, and useful. The team is focused on infrastructure reliability, operating GPU clusters and networks for AI workloads.

Latin America

  • Design and own the playbook for AI productization, governance, and reusable internal tooling.
  • Champion AI-assisted development, TDD, and modern CI/CD practices across the organization.
  • Act as the primary technical authority bridging AI/ML architecture and executive strategy.

RYZ Labs is a startup studio that creates industry-defining companies by leveraging top talent and cutting-edge cloud technologies. Founded in 2021, we have a remote team across the US and Latam, with a culture of autonomy, ownership, and continuous improvement.

EMEA

  • Serve as the strategic technical thought partner to EMEA partner leadership, owning the technical relationship and setting the vision for how Coder fits into each partner's AI and modernization portfolio.
  • Co-create sales plays, develop workshops and Builder Labs, and intervene directly on high-value partner-attached deals across EMEA to drive pipeline and technical wins.
  • Build repeatable assets like reference architectures, enablement modules, and GTM playbooks that scale across the ecosystem without requiring your direct involvement.

Coder is an AI software development company leading the future of autonomous coding by empowering teams to build software faster, more securely, and at scale through AI coding agents and human developers. As a growing startup, we foster a culture of autonomy, innovation, and collaboration, focused on making agentic AI a safe and integral part of every software development lifecycle.

Global

  • Own end-to-end delivery quality for major engagements, translating ambiguous client needs into practical execution plans.
  • Lead solution architecture and technical decision-making, making pragmatic tradeoffs between speed, quality, and client value.
  • Build and ship production AI/ML systems using Python, ML frameworks, and cloud-native infrastructure while mentoring other engineers.

Eliza is a technology services company and Advanced-tier OpenAI partner that helps organizations build and deploy AI solutions, from generative AI to predictive analytics. They are a collaborative, mission-driven team focused on real-world AI impact.

Global 6w PTO 26w maternity 26w paternity

  • Lead end-to-end deployment of North in private cloud and on-premises environments, including planning, configuration, testing, and rollout.
  • Partner with enterprise IT teams to assess infrastructure, security requirements, and data management practices.
  • Design and implement deployment strategies tailored to client needs, ensuring compliance with data privacy and security standards.

Cohere is a security-first enterprise AI company that builds cutting-edge foundation models and end-to-end products. It is a global team of researchers, engineers, and designers passionate about their craft.

North America Unlimited PTO

  • Deploy and scale MCP-based AI agents on Kubernetes for enterprise customers across the US-East and EMEA regions.
  • Lead complex technical engagements, build reusable deployment patterns, and mentor engineers on the team.
  • Shape product roadmap by feeding back field insights from regulated industries and defining regional engagement standards.

Stacklok builds the control plane for enterprise AI agents, enabling organizations to run, govern, and secure them on Kubernetes and private cloud. Founded by two Kubernetes creators, the company is already adopted by leading tech and regulated industries, fostering a collaborative, AI-maximalist culture with deep open-source roots.

Switzerland

  • Design and build production-grade ML inference infrastructure using frameworks like vLLM and Triton.
  • Optimize GPU utilization, memory efficiency, and model artifact storage for cost-effective performance.
  • Collaborate with infrastructure and AI teams to establish engineering best practices and scalable platform architecture.

Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. It uses technology to review applications and share top candidate shortlists with employers, operating in a remote-first environment.

$152,100–$190,100/yr
US

  • Identify and map workflows to find step-change opportunities for AI automation.
  • Design and build future-state workflows using agents, integrations, and human-in-the-loop checkpoints.
  • Deploy and run agents in production, tracking KPIs and iterating on performance.

Natera is a global leader in cell-free DNA testing, focusing on oncology, women's health, and organ health. The company employs a diverse team of dedicated professionals from world-class institutions, fostering a collaborative and inclusive culture.

$225,000–$300,000/yr
Unlimited PTO

  • Own software architecture and delivery: design, deploy, and operate systems at scale.
  • Design and evolve AI systems including prompting, retrieval, evaluation infrastructure, and expert feedback loops.
  • Hire, manage, and grow the engineering team from 3 FTEs while staying on the frontier of AI and legal AI.

Inhouse is the #1 AI lawyer for small to midsize businesses, combining AI, their own law firm, and an expert feedback loop to deliver fast, compliant legal work. They grew revenue 1,500% last year and recently raised a $5M seed round from leading VCs.

North America

  • Seeking an early-stage operator experienced in AI infrastructure and compute platforms to build a company in the AI compute space.
  • The ideal candidate has run ML platform or infrastructure teams deploying inference across cloud, neocloud, and owned GPU/edge hardware.
  • Must be Canada or US-based and able to collaborate in North American time zones.

Forum Ventures is a venture studio that builds B2B SaaS businesses from 0 to 1, providing capital, networks, and resources to founders. They have launched 17 companies since 2023 and are launching 7 more in 2026, with a culture focused on speed, insight, and success.

UK

  • Design, build, and operate Kubernetes infrastructure for AI workloads using Terraform and GitOps.
  • Define SLOs, run incident response, and create runbooks for reliable AI platform operations.
  • Drive AI-specific observability, FinOps, and security practices across the platform.

We are an AI-native consulting partner working with clients like PayPal, adidas, and NatWest to build digital products and services. Our team of over 600 has scaled quickly, earning Great Place to Work-Certified status multiple years in a row.

US

  • Monitor, operate, and support production AI infrastructure platforms and resolve incidents.
  • Investigate performance, availability, and reliability issues across infrastructure and platform components.
  • Collaborate with engineering teams, hardware vendors, and data center personnel to resolve technical issues.

Mirantis is the Kubernetes-native AI infrastructure company, enabling organizations to build and operate scalable, secure, and sovereign infrastructure for modern AI, machine learning, and data-intensive applications. It is a Silicon Valley leader with passionate, talented colleagues, offering a competitive compensation package and strong benefits.

Europe

  • Lead technical operations for large-scale AI infrastructure environments powered by NVIDIA GPUs and Kubernetes.
  • Act as a senior escalation point for critical incidents and drive root cause analysis and long-term corrective actions.
  • Mentor team members and shape operational standards, automation, and reliability practices for next-generation platform services.

Mirantis is the Kubernetes-native AI infrastructure company, enabling organizations to build and operate scalable, secure, and sovereign infrastructure for modern AI, machine learning, and data-intensive applications. Mirantis serves many of the world’s leading enterprises, including Adobe, DocuSign, Liberty Mutual, PayPal, Reliance Jio, Societe Generale, Splunk, and Volkswagen.

$185,500–$265,000/yr
US

  • Design, evolve, and scale the Customer Success AI platform, including reusable services and reference architectures that support multiple AI products.
  • Build and maintain foundational AI infrastructure such as orchestration services, context engines, memory services, and workflow execution frameworks.
  • Optimize platform capabilities across model routing, latency, cost, reliability, observability, governance, and security for enterprise-grade performance.

Zscaler accelerates digital transformation by providing a cloud-native Zero Trust Exchange platform that secures users, devices, and applications from cyberattacks and data loss. The company is building a culture of execution centered on customer obsession, collaboration, ownership, and accountability with high-performing teams.

US

  • Evolving the AI knowledge platform with retrieval, indexing, and synthesis for organization-wide use.
  • Architecting and operating agentic infrastructure on AWS with cost guardrails and observability.
  • Partnering with product engineering to define the AI platform API surface and building reference agent implementations.

ShiftKey is a healthcare workforce marketplace that connects facilities with licensed professionals to fill shifts, addressing staffing shortages. The company fosters an inclusive and collaborative culture, valuing diverse perspectives.