Collaborate with Data Science, Product Managers, and Software Engineers to build robust ETL pipelines for user-facing features.
Contribute to architecture decisions, observability tooling, and data quality initiatives to keep the platform robust.
Enforce engineering best practices across the AI/ML org, including code quality, testing, and documentation.
Federato is an AI-native platform for insurance, enabling insurers to provide affordable coverage for climate, cyber, and social inflation risks. It is a small, well-funded company backed by the investors behind Salesforce, Veeva, and Zoom, with a culture focused on first principles, learning, and fun.
Define data architecture and platform strategy for enterprise data platforms.
Build and optimize scalable data pipelines supporting batch and real-time processing.
Mentor junior engineers and partner with leadership on data strategy.
Robots & Pencils is an applied AI engineering firm building AI co-workers for enterprise operations. Founded in 2009, with delivery centers across North America and Europe, our teams average 15+ years of experience and we value craft, speed, and ownership.
Design, develop, and maintain resilient, scalable, end-to-end data pipelines with Snowflake and Python.
Partner with product owners and analysts to gather requirements and deliver data solutions aligned with business goals.
Act as a subject matter expert, mentor junior engineers, and champion DevOps practices including CI/CD and automated testing.
Privia Health is a technology-driven national physician enablement company that collaborates with medical groups and health systems to optimize practices and improve patient care through its cloud-based platform. The company is led by top industry talent and physician leadership, focusing on reducing healthcare costs and improving outcomes.
Design, build, and own data pipelines moving data from application and third-party sources into relational databases.
Build and ship production AI systems, including data infrastructure and evaluation systems for feature reliability.
Set standards for data work through clear writing, early context sharing, and team efficiency.
Dscout builds a flexible UX research platform trusted by top brands in finance, healthcare, and tech. They are a remote-first team of passionate professionals that prioritize learning, diversity, and inclusion.
Build advanced data pipelines using Snowflake and AWS, leveraging AI-assisted development and Medallion Architecture.
Design analytical data models and write complex Python/SQL to ensure high performance and 99.95% uptime.
Mentor junior engineers and drive architectural improvements, supporting data governance and automation.
Spring Venture Group is a digital direct-to-consumer sales and marketing company specializing in Medicare Supplement and related products. It is a leading company with a family of brands and a dedicated team of licensed insurance agents, fostering a diverse and inclusive culture.
Design, build, and operate scalable Data & AI platform capabilities for analytics, data science, and Generative AI use cases.
Develop reusable frameworks, services, and workflows to standardize how teams build and manage data pipelines and data products.
Partner with Analytics, Applied-AI, and Engineering teams to deliver platform capabilities that support business and clinical decision-making.
Omada Health is on a mission to bend the curve of chronic disease by providing a virtual-first care model that combines human-led care teams, connected devices, and AI-enabled technology. They have served over two million members and are a publicly traded company with a culture of trust, context, boldness, results, and teamwork.
Build and maintain data pipelines for analytics, ML, and product applications.
Design scalable data infrastructure with a focus on quality and observability.
Collaborate with cross-functional teams to understand data needs and implement solutions.
Prolific builds human data infrastructure to power the next wave of AI innovation. They are a remote-first company focused on ethical data collection and mission-driven culture.
Design, build, and maintain data pipelines that ingest and transform healthcare data through staging, transformation, and mart layers.
Write and maintain dbt models following established conventions, ensuring transformations are testable and performant.
Implement data quality tests, monitor pipeline health, and resolve data anomalies to ensure accuracy of downstream reporting.
Privia Health is a technology-driven physician enablement company that partners with medical groups and health systems to optimize practices and improve patient care. The company is led by top industry talent and uses cloud-based technology to reduce healthcare costs and improve outcomes.
Design and build production data pipelines using Databricks technologies such as Lakeflow, Autoloader, and Structured Streaming.
Architect and implement Lakehouse solutions on Databricks with medallion architecture, Delta Lake, and Unity Catalog.
Consult with clients to understand data challenges and implement sustainable data strategies and solutions.
Livefront helps companies design and build world-class digital products that command attention and inspire joy. They are a consultancy with a reputation for excellence, working with household names and startups, and they value collaboration, quality, and community.
Design and implement robust, scalable data ingestion and transformation pipelines using Databricks, PySpark, and distributed processing.
Implement Delta Lake principles focusing on CDC and schema evolution, and integrate data quality frameworks within CI/CD pipelines.
Develop and optimize complex SQL and Python scripts, handling diverse data sources and supporting data governance solutions.
Mobile Wave Solutions is a professional services company specializing in software development as a service. With a team of over 120 engineers, we deliver scalable, high-quality software that empowers our global clients to innovate and grow.
Design, build, and maintain robust, scalable ELT/ETL data pipelines from various source systems into cloud data platforms.
Perform data modeling, including dimensional modeling, and build transformation layers using dbt to create analytics-ready datasets.
Support operational reliability, monitor data pipelines, and ensure SLAs for timeliness, freshness, and accuracy.
Troveo builds the data platform that AI labs and model builders need to train the next generation of models. Backed by top investors, we’re a small, high-impact team solving one of the biggest bottlenecks in AI development.
Build a cloud-native big data platform handling audience data for millions of attendees and billions of interactions.
Design and own ML infrastructure including feature stores, training pipelines, and model serving.
Own the full data pipeline from ingestion to business impact, ensuring reliability and performance.
Hive is a marketing platform that helps event marketers personalize and automate campaigns to sell out shows and engage fans. The company is a fully remote team located across Canada, fostering a work environment with strong work-life balance and a focus on impactful outcomes.
Lead the Data Engineering team to design, build, and operate scalable data pipelines powering analytics and AI.
Partner with cross-functional leaders to define technical strategy, roadmap, and execution for the enterprise data platform.
Drive engineering excellence through improved data quality, reliability, observability, and governance.
Omada Health is reverse engineering healthcare delivery in America, focusing on chronic conditions like obesity, diabetes, and hypertension. With over two million members served and a strong remote-first culture, the company has been certified as a Great Place to Work.
Work cross-functionally with Product and experts to conceptualize, prototype, and build data solutions
Build and maintain data engineering systems and high-quality data models from multi-source healthcare datasets
Develop and test data pipelines and draft internal and external technical documentation
Turquoise Health is a Series C price transparency platform for finance leaders across healthcare. Backed by a16z, Oak HC/FT, and others, we are a remote-first, US-based team that values transparency, empathy, inclusivity, creativity, and ownership.
Design, build, and maintain scalable data and ML pipelines for analytics and AI systems.
Build and optimize workflows for structured and unstructured data, enabling semantic search and RAG use cases.
Manage and optimize vector databases and indexing strategies for efficient retrieval and AI-powered search.
This is a partner company seeking a Data & Machine Learning Engineer based in Brazil. They operate in a highly technical and global environment with strong emphasis on scalability, performance, and innovation.
Lead, coach, and develop a global team of data engineers while guiding technical architecture and delivery.
Build and scale large-scale data pipelines, production datasets, and QA systems using Databricks, Airflow, SQL, and PySpark.
Partner with product and research teams to transform complex data into reliable, production-grade assets for customers and internal applications.
YipitData is a leading market research and analytics firm that analyzes billions of alternative data points to provide actionable insights for the disruptive economy. The company recently raised $475M, operates globally with offices in the US, APAC, and India, and has been recognized by Inc. as a Best Workplace for three consecutive years.
Builds and maintains feature pipelines, training datasets, and forecast workflows for revenue, demand, and operational planning.
Operationalizes forecasting and ML models through repeatable training, evaluation, deployment, and monitoring in Databricks.
Implements model performance tracking, drift monitoring, and traceability from source data to prediction output.
Sonny's Enterprises is the world's largest manufacturer of conveyorized car wash equipment, parts, and supplies. The company is the industry leader with a culture focused on innovation and growing its people.
Own Databricks production support for the Sugar Predict data platform, including monitoring and incident response.
Migrate legacy ETL pipelines to Databricks, building automation tooling to reduce manual intervention.
Design and build high-performance Databricks pipelines that ingest and serve ERP and CRM data at scale across Azure and AWS.
SugarAI is redefining CRM for the age of AI by turning fragmented customer and revenue signals into clear, prioritized action. With a global team committed to impact, ownership, and continuous growth, we create an environment where thoughtful ideas move quickly and people are trusted to lead.
Build and scale reliable data infrastructure powering analytics, AI-driven decision-making, and smarter growth across n8n.
Own critical parts of the modern data stack, including end-to-end data pipelines using dbt on BigQuery and orchestration with Dagster.
Enable analytics, AI use cases, and marketing attribution by delivering trusted datasets and scalable data products.
n8n is an open workflow orchestration platform built for the new era of AI, giving technical teams the freedom of code with the speed of no-code. The company has grown to a team of over 260 across Europe and the US, backed by top investors, with a culture of builder spirit and remote-first collaboration.
Design, build, and maintain data infrastructure and pipelines spanning both batch and real-time workloads.
Build and maintain scalable ETL/ELT pipelines using Python, SQL, Spark, and orchestration frameworks.
Partner with data scientists, ML engineers, and product teams to deliver data products and establish data quality frameworks.
Gemini is a global crypto and Web3 platform founded by Cameron and Tyler Winklevoss in 2014, offering a wide range of crypto products to individuals and institutions in over 70 countries. As a publicly traded company, Gemini is poised to accelerate the vision of reshaping the global financial system with greater scale, reach, and impact.