Build and own the reusable on-device ML inference platform for Life360's portfolio of trackers and wearables, integrating edge ML into resource-constrained RTOS firmware.
Perform core firmware engineering including low-level driver development, power management, and debugging on real hardware, while carrying regular firmware work when ML demand is light.
Develop and ship ML models on device, handling quantization, optimization, and the full pipeline from sensor data to inference result, and set technical direction for on-device intelligence.
Life360's mission is to keep people close to the ones they love, offering a mobile app, Tile tracking devices, and Pet GPS tracker for location sharing, safe driver reports, and crash detection. The company serves approximately 97.8 million monthly active users across over 180 countries and has more than 500 remote-first employees, fostering a collaborative and mission-driven culture.
Build and improve ML/CV models, including retraining and optimizing for Samsara-specific problems.
Work with petabyte-scale data from camera and sensor devices to develop new models.
Partner with firmware and full-stack teams to deploy models for optimal performance and cost.
Samsara is the pioneer of the Connected Operations Cloud, enabling organizations to harness IoT data for actionable insights to improve safety, efficiency, and sustainability. As a public company with over 2.3 million IoT devices deployed globally, Samsara fosters a culture of growth mindset, inclusivity, and teamwork.
Deploy LLMs into production across GPU infrastructure, owning the full pipeline from customer query to served response.
Stand up and operate serving infrastructure using vLLM, SGLang, or TensorRT-LLM.
Apply quantization, batching, caching, and routing to optimize latency and cost at scale.
vCluster Labs is a venture-backed tech startup pioneering Kubernetes virtualization for the AI era, enabling AI Cloud providers and AI factories to operate GPU infrastructure with hyperscaler-like experiences. We raised over $30M from top-tier VCs like Khosla Ventures, are in a hyper-growth phase, and maintain a remote-first, distributed global team with headquarters in San Francisco.
Design and implement cloud/edge AI architectures for real-time computer vision applications.
Develop computer vision models for wildfire smoke detection, vegetation classification, asset detection, and spatial reasoning.
Port and optimize deep learning models for ARM64, CUDA, TensorRT, ONNX, and NVIDIA Jetson platforms.
Pano AI is a leader in AI-powered wildfire detection and intelligence, helping fire professionals detect, respond to, and contain wildfires faster. Our team of 175+ works in a hybrid-remote environment across North America and Australia, based in San Francisco.
Architect end-to-end computer vision pipelines for real-world safety detection, including object detection, tracking, and semantic segmentation.
Drive the technical roadmap for edge and cloud perception, optimizing models for constrained hardware without sacrificing accuracy.
Mentor senior scientists and engineers, translating customer needs into precise engineering challenges.
Samsara is the pioneer of the Connected Operations™ Cloud, a platform that enables organizations dependent on physical operations to harness IoT data for actionable insights and improved safety, efficiency, and sustainability. As a recently public company, they foster a culture of rapid career development, collaboration, and innovation while scaling globally.
Design and maintain ML model productionization infrastructure for high-visibility product features.
Collaborate with data science to streamline model training, validation, and deployment.
Implement robust monitoring and alerting for model performance, drift, and data quality.
The Athletic is a sports media company powered by one of the largest global newsrooms in sports, delivering in-depth coverage of professional and college teams across North America and Europe. With over 500 full-time staff, they foster a collaborative culture focused on high-quality journalism and data-driven innovation.
You will explore opportunities in edge AI, focusing on runtime abstraction, inference routing, and unified operational planes for heterogeneous hardware.
You will leverage a $250K investment, fundraising playbook, and operating team support to validate your idea and achieve early traction.
You should have early-stage operator experience in edge ML and on-device inference, with a passion for building venture-backed companies.
Forum Ventures is a venture studio that brings together ambitious people, brilliant ideas, and capital to build B2B SaaS businesses from 0 to 1. They have launched 17 companies since 2023 and are launching 7 more in 2026, with a focus on helping founders move faster and build successful startups.
Build and own the model serving infrastructure, real-time inference, feature retrieval, and the latency budget that governs both.
Build the deployment path for data scientists to ship models, including bring-your-own-model support.
Own models in production: monitoring, drift detection, retraining, incident response, and the on-call rotation.
Sardine is the leading agentic risk platform for fighting financial crime. We are a remote-first company with hubs in the Bay Area, NYC, Austin, Toronto, and São Paulo, hiring talented individuals with extreme ownership and high growth orientation.
Assist in developing computer vision models for wildfire detection and environmental monitoring.
Help implement and maintain ML/CV pipelines and deploy models on NVIDIA Jetson edge platforms.
Collaborate with AI researchers, software engineers, and product teams to optimize and document solutions.
Pano AI is the leader in AI-powered wildfire detection and intelligence, helping fire professionals detect, respond to, and contain wildfires faster and more safely. We are a team of more than 175 people working in a hybrid-remote environment across North America and Australia, with headquarters in San Francisco.
Implement and optimize compute kernels for Attention, GEMM, MoE, and quantization on NVIDIA, AMD, or AWS Trainium using CUDA, Triton, ROCm/HIP, or Neuron SDK.
Profile and improve inference performance in vLLM, SGLang, and custom runtimes through kernel fusion, scheduling, and memory optimizations.
Ship code upstream to open-source AI infrastructure projects with tests and documentation, working on a well-scoped project from design to production.
Yotta Labs is building the next generation multi-silicon AI cloud and runtime platform to power the world’s most demanding AI workloads. They are a remote-first team with a focus on high-performance AI computing and offer a flexible, collaborative work environment.
Build the software backbone for autonomous foundation models, including multimodal data pipelines and training workflows.
Implement and iterate on LLM, VLM, and VLA architectures, owning model code paths and inference runners.
Deliver production-grade serving tooling for low-latency operation and own systems from architecture to iteration.
We're building the next generation of ground transportation with advanced physical AI to simplify freight challenges. Our stealth team, founded by engineers who scaled autonomous driving, is developing a new vehicle platform and focuses on creating reliable real-time autonomous systems.
Build and deploy production code to support customer AI inference workloads on Tenstorrent's hardware and software stack.
Debug and optimize across the full inference stack, from serving layer to kernel dispatch, and translate customer issues into actionable requirements.
Operate Kubernetes and observability tools to manage multi-node AI clusters and ensure reliability.
Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. Their diverse team of technologists has developed a high-performance RISC-V CPU from scratch, and they value collaboration, curiosity, and a commitment to solving hard problems.
Independently own high-value optimization initiatives across training, inference, or launch-readiness for important Ads ML workloads.
Diagnose bottlenecks in real production systems using profiling, benchmarking, and observability.
Build performance tooling, optimization playbooks, and efficiency primitives that benefit multiple teams.
Reddit is a community of communities built on shared interests and authentic conversations. With 100,000+ active communities and approximately 126 million daily active unique visitors, Reddit has a flexible workforce and values collaboration.
Design and maintain CI/CD and MLOps pipelines for AI and software applications, ensuring seamless deployment and automation.
Build and scale cloud-native infrastructure using Kubernetes, Docker, and GPU clusters to support high-performance AI workloads.
Champion Infrastructure as Code and observability practices to ensure high availability, security, and compliance across multi-cloud environments.
Bitdeer is a world-leading technology company providing AI and Bitcoin mining infrastructure. Headquartered in Singapore, the company has a global presence with data centers in multiple countries and a culture focused on innovation and reliability.
Lead technical operations for large-scale AI infrastructure environments powered by NVIDIA GPUs and Kubernetes.
Act as a senior escalation point for critical incidents and drive root cause analysis and long-term corrective actions.
Mentor team members and shape operational standards, automation, and reliability practices for next-generation platform services.
Mirantis is the Kubernetes-native AI infrastructure company, enabling organizations to build and operate scalable, secure, and sovereign infrastructure for modern AI, machine learning, and data-intensive applications. Mirantis serves many of the world’s leading enterprises, including Adobe, DocuSign, Liberty Mutual, PayPal, Reliance Jio, Societe Generale, Splunk, and Volkswagen.
Serve as the primary technical point of contact for teams running large-scale training and inference workloads, owning onboarding end to end.
Diagnose and resolve complex failures in customer environments, from network fabric to ML frameworks, and build automation to prevent recurrence.
Profile and improve distributed training performance, lead incident response, and turn field insights into product improvements.
Andromeda Cluster provides scaled AI infrastructure to early-stage startups, founded by Nat Friedman and Daniel Gross. They work with leading AI labs, data centers, and cloud providers to deliver compute globally, building the liquidity layer for AI compute.
Optimize machine learning inference systems for latency, throughput, and cost-efficiency.
Profile and troubleshoot GPU/CPU bottlenecks, implement advanced techniques like quantization and speculative decoding.
Collaborate with research and engineering teams to productionize new models and improve inference infrastructure.
The company is an AI-focused organization that develops advanced machine learning systems for production environments. It values technical excellence and experimentation, offering a flexible remote work environment.
Monitor, operate, and support production AI infrastructure platforms including NVIDIA GPU environments.
Investigate and resolve infrastructure, networking, hardware, and platform-related incidents.
Collaborate with engineering teams, vendors, and datacenter personnel to improve operational processes.
Mirantis is the Kubernetes-native AI infrastructure company, enabling organizations to build scalable, secure, and sovereign infrastructure for modern AI and data-intensive applications. Serving enterprises like Adobe and PayPal, the company combines open source innovation with deep Kubernetes expertise to deliver composable developer platforms across any environment.
Design, implement, and operate production-quality infrastructure with significant autonomy.
Build automated testing infrastructure for software running across heterogeneous edge hardware.
Improve CI/CD pipelines, build and release systems, deployment automation, developer tooling, and platform reliability.
TurbineOne is a frontline perception company delivering decision advantage and situational awareness for military intelligence. It is a fast-moving, high-performance startup with a small, fully remote team where ownership matters.
Design and implement scalable real-time data integration and Change Data Capture (CDC) solutions using the Striim platform.
Build proof-of-concepts, reference architectures, and deployment patterns for enterprise implementations.
Collaborate with Engineering, Product, and GTM teams to validate architectural designs and improve platform capabilities.
Striim is a unified data integration and streaming platform that connects clouds, data, and applications with real-time analytics for enterprise customers. It is a Silicon Valley startup with a culture fostering entrepreneurship and growth, operating as one team with unlimited potential and dignity.