Build and own the reusable on-device ML inference platform for Life360's portfolio of trackers and wearables, integrating edge ML into resource-constrained RTOS firmware.
Perform core firmware engineering including low-level driver development, power management, and debugging on real hardware, while carrying regular firmware work when ML demand is light.
Develop and ship ML models on device, handling quantization, optimization, and the full pipeline from sensor data to inference result, and set technical direction for on-device intelligence.
You will explore opportunities in edge AI, focusing on runtime abstraction, inference routing, and unified operational planes for heterogeneous hardware.
You will leverage a $250K investment, fundraising playbook, and operating team support to validate your idea and achieve early traction.
You should have early-stage operator experience in edge ML and on-device inference, with a passion for building venture-backed companies.
Forum Ventures is a venture studio that brings together ambitious people, brilliant ideas, and capital to build B2B SaaS businesses from 0 to 1. They have launched 17 companies since 2023 and are launching 7 more in 2026, with a focus on helping founders move faster and build successful startups.
Coordinate with AI and Network Software Engineers to design low-latency multi-node communication frameworks in restricted and/or disconnected environments.
Quantize, compile, and deploy real-time multi-agent pipelines for readiness across diverse edge architectures.
Develop algorithms for asynchronous multimodal data fusion, tracking, and signal processing.
webAI is pioneering the future of artificial intelligence by establishing the first distributed AI infrastructure for personalized AI, focusing on private deployment at the edge. The company is a team driven by truth, ownership, tenacity, and humility, seeking individuals passionate about shaping next-generation AI.
Own the entire platform layer: operating system, firmware, streaming stack, and releases across all Airtame devices.
Spend roughly 60% of time on the hardest technical problems and 40% growing the engineers on the team.
Drive key outcomes: take AirPlay and Miracast from beta to default, make firmware releases routine, and lead new hardware bringup.
Airtame develops hardware and software for wireless screen sharing, enabling seamless connectivity in meeting rooms. The company has around 55 employees, headquartered in Copenhagen, with a cross-functional, hybrid team that values work-life balance and open communication.
Develop and optimize low-level kernels, runtime components, and system software for high-performance AI inference workloads.
Improve inference engine performance across GPU platforms by identifying bottlenecks and implementing advanced optimization techniques.
Profile, debug, and resolve system-level and hardware-level performance issues across CPU and GPU environments.
This position is listed on behalf of a partner company building cutting-edge AI infrastructure for large-scale inference platforms. They operate in a highly technical, international, and innovation-driven environment where engineering excellence and ownership are valued.
Lead the design and development of our production inference platform, defining the technical roadmap for inference infrastructure, model serving, and runtime optimization.
Build and operate scalable, cost-effective systems for serving large language models in production, optimizing latency, throughput, GPU utilization, and memory efficiency.
Partner with ML engineers to productionize new models and inference techniques, establish benchmarking methodologies, and make key architectural decisions.
Syllo is on a mission to transform litigation with a unified platform that enables lawyers to safely harness AI. Since going to market, they have gained diverse enterprise customers including big law firms and corporations, and are quickly expanding.
Drive low-level Qualcomm Android BSP bring-up, devicetree modifications, and pSIM framework implementations for solar-powered LPR and live-video cameras.
Collaborate across electrical engineering and application teams to bridge hardware architectures with cloud-connected software for life-saving devices.
Eliminate system vulnerabilities and accelerate hardware execution to ensure our devices never drop a frame in demanding real-world conditions.
Flock builds technology that reduces crime and protects privacy, partnering with cities, businesses, schools, and neighborhoods. They are a high-performance team with over $1B in funding and an $8.3B valuation, scaling with intention and a culture of urgency and ownership.
Define what to build and how to build it, shaping product direction from the engineering side.
Own the technical direction for orchestrating concurrent, real-time interactions to deliver one coherent experience for the driver.
Build and ship new AI-powered products from zero to one in a fast-moving, ambiguous environment.
Samsara is the pioneer of the Connected Operations Cloud, helping organizations that depend on physical operations harness IoT data to improve safety, efficiency, and sustainability. As a recently public company, they offer autonomy and support to make an impact with a culture focused on long-term growth and customer success.
Implement and optimize compute kernels for Attention, GEMM, MoE, and quantization on NVIDIA, AMD, or AWS Trainium using CUDA, Triton, ROCm/HIP, or Neuron SDK.
Profile and improve inference performance in vLLM, SGLang, and custom runtimes through kernel fusion, scheduling, and memory optimizations.
Ship code upstream to open-source AI infrastructure projects with tests and documentation, working on a well-scoped project from design to production.
Yotta Labs is building the next generation multi-silicon AI cloud and runtime platform to power the world’s most demanding AI workloads. They are a remote-first team with a focus on high-performance AI computing and offer a flexible, collaborative work environment.
Partner directly with customers to architect, build, and deploy high-scale production AI applications on Baseten's platform.
Own the journey from initial exploration to production deployment, translating ambiguous goals into reliable services.
Work across product, software development, performance engineering, and customer-facing implementations.
Baseten powers mission-critical inference for dynamic AI companies like Cursor and Notion. They are rapidly growing, recently raised a $1.5B Series F, and foster a collaborative, forward-thinking culture.
Design and implement AI capabilities for intelligent data characterization and decision support.
Evaluate, optimize, and deploy open-weight foundation models for resource-constrained edge environments.
Develop efficient inference pipelines and implement RAG, semantic search, and model optimization techniques.
Expression provides data fusion, analytics, AI/ML, software engineering, and spectrum management solutions to the U.S. Department of Defense and national security community. Founded in 1997 and headquartered in Washington DC, the company was ranked #1 on Washington Technology's 2018 Fast 50 and is a Top 20 Big Data Solutions Provider, fostering a collaborative culture with opportunities for growth.
Own end-to-end Machine Learning (ML) system execution including data pipelines, training, and deployment.
Fine-tune and adapt models using state-of-the-art methods like LoRA and DPO.
Architect scalable inference systems and collaborate closely with application engineering.
This company develops advanced production-grade machine learning systems. The team is small and high-trust, with a culture of ownership and pragmatism.
Write and modify embedded C firmware for analog and Ethernet connectivity devices, including register configuration, state machines, and calibration sequences.
Develop Python scripts to automate test flows, collect and analyze data, and help build CI infrastructure.
Participate in code reviews, architecture discussions, and maintain project tracking with Jira.
Astera Labs provides rack-scale AI infrastructure through purpose-built connectivity solutions. The company is a publicly-traded startup with a small, fast-moving team that values collaboration and ownership.
Optimize machine learning inference systems for latency, throughput, and cost-efficiency.
Profile and troubleshoot GPU/CPU bottlenecks, implement advanced techniques like quantization and speculative decoding.
Collaborate with research and engineering teams to productionize new models and improve inference infrastructure.
The company is an AI-focused organization that develops advanced machine learning systems for production environments. It values technical excellence and experimentation, offering a flexible remote work environment.
Independently own high-value optimization initiatives across training, inference, or launch-readiness for important Ads ML workloads.
Diagnose bottlenecks in real production systems using profiling, benchmarking, and observability.
Build performance tooling, optimization playbooks, and efficiency primitives that benefit multiple teams.
Reddit is a community of communities built on shared interests and authentic conversations. With 100,000+ active communities and approximately 126 million daily active unique visitors, Reddit has a flexible workforce and values collaboration.
Lead a team of platform engineers to build and operate ML training and serving infrastructure, including GPU and low-latency serving.
Combine strong people leadership with technical judgment in ML infrastructure, partnering with senior ICs and cross-functional teams.
Drive delivery, operational health, and evaluate modern ML tooling to support company-wide ML priorities.
Affirm is reinventing credit to make it more honest and friendly, offering buy now pay later solutions without hidden fees. It is a remote-first company with a strong engineering culture, prioritizing people and providing competitive benefits.
Architect, develop, debug, and test customer facing firmware features written in C/C++ and Golang for Samsara’s cloud-connected hardware platform.
Build IOT products for the long term and support our 1.5+ million devices in the field.
Design and develop our internal observability discipline to ensure the health of our firmware rollouts.
Samsara is the pioneer of the Connected Operations Cloud, enabling organizations to harness IoT data to improve safety, efficiency, and sustainability. With over 2.3 million IoT devices deployed, it is a recently public company that values autonomy and support for long-term impact.
Develop software across the Mission Solutions portfolio using modern AI tools.
Work alongside a senior Solutions Architect to take ownership of features.
Support field and integration activities with up to 25% travel.
Ultra Intelligence & Communications specializes in application-engineered bespoke solutions for mission-critical systems in defense, security, and detection markets. The company employs over 4,500 employees globally and emphasizes a collaborative, employee-focused culture.
Define and drive innovative technical vision for intelligent networking and agentic AI platforms.
Design and build high-performance, production-ready services for real-time data and AI processing.
Mentor engineers, lead technical discussions, and uphold engineering excellence across the team.
We provide end-to-end, cloud-driven networking solutions trusted by over 50,000 customers globally. With double-digit growth and a culture of inclusion, we foster an innovative workplace where all employees thrive.
Bridge the gap between market and internal teams by translating client feedback into actionable product improvements.
Build deep technical relationships with customer architects and serve as their primary technical advisor.
Train customers on implementing Spiking Neural Processor technology through workshops, seminars, and on-site support.
Innatera is a rapidly growing Dutch semiconductor company that develops ultra-efficient neuromorphic processors for AI at the edge. The company fosters a dynamic, fearless engineering culture with ambitious teams and an inclusive environment.
Own end-to-end ML system execution including data pipelines, training workflows, evaluation systems, inference architecture, and deployment.
Fine-tune and adapt models using state-of-the-art methods such as LoRA, QLoRA, SFT, DPO, and distillation.
Architect scalable inference systems, balance latency, cost, and reliability, and deploy production-grade ML solutions.
Gina's Tech Jobs is a recruiting and staffing company that helps firms hire technical talent. They are a small agency focused on IT roles, fostering a high-trust, collaborative environment.