Design and build the control plane for our curated data lifecycle, tackling orchestration, backfills, restatements, retries, and partial failure recovery.
Decide dataset-by-dataset whether the answer is a model, service, or job, and own that architecture through production.
Work across Go, Kotlin, Rust, Python, and SQL, choosing the right tool rather than the familiar one.
Turn product requirements into technical requirements, model the domain, define interfaces, and navigate trade-offs.
Break large problems into work others can own and sequence delivery for early value.
Drive high-impact architectural decisions across the data platform and own components end to end.
Dune is a collaborative multi-chain analytics platform that makes crypto data accessible to developers, analysts, and investors. The team of ~60 employees works across Europe and eastern US timezones with a fully remote-first approach.
Design and build distributed data systems handling large-scale ingestion and processing.
Drive architectural decisions and take end-to-end ownership of critical components.
Collaborate with product teams to translate ambiguous requirements into robust technical solutions.
Our partner builds a large-scale, multi-chain data platform that ingests, models, and delivers blockchain data to users and developers. They are a remote-first, distributed team with a strong engineering culture focused on ownership and collaboration.
Build and maintain data ingestion and transformation pipelines using Dagster, dbt, and Snowflake through PR-driven development.
Manage production monitoring, backfills, and Airbyte connections to ensure reliable daily data flows.
Collaborate with analytics and product teams on dimensional modeling and data quality, and participate in code review and runbook writing.
PadSplit is at the forefront of solving the affordable housing crisis through its tech-driven platform. They have a growing remote-first team and offer a competitive benefits package including unlimited PTO and paid parental leave.
Build and maintain ELT pipelines using Snowflake and dbt to power the company's data platform.
Own the reliability and performance of the data stack, including testing, monitoring, and alerting.
Collaborate with internal teams and external clients to translate business needs into durable data solutions.
Jump is transforming the live sports experience with an end-to-end fan engagement platform for sports teams and venues. They are a remote-first team backed by top investors, driven by values of trust, teamwork, and innovation.
Build and maintain BigQuery data models using Dataform, following medallion architecture patterns (Bronze/Silver/Gold).
Contribute to Looker dashboards and LookML models, working alongside senior engineers and analysts.
Build and maintain robust Python data pipelines with testing, linting, and CI/CD integration.
Kitman Labs is a performance intelligence company that transforms how the sports industry uses data to unlock athlete potential. With over 2000 teams in 50 leagues across 6 continents, the company has assembled a team of top data scientists, sports performance scientists, and product specialists, fostering an innovative and collaborative culture.
Design and build the data systems that power analytics, AI, and business intelligence across Owner
Replace one-off custom scripts with reproducible, production-grade pipelines
Own and evolve our core platforms - Snowflake, dbt, Dagster, and our ingestion and reverse ETL layers
Owner is an AI-native platform that replaces multiple tools for local business owners, starting with restaurants. The company has a low-hundreds team with top talent from successful SMB software companies and a remote-first global culture.
Design and build scalable backend services and APIs for LaunchDarkly's Experimentation product.
Own backend components from discovery through monitoring and continuous improvement.
Collaborate with cross-functional teams to translate customer needs into reliable technical solutions.
LaunchDarkly is a feature management platform that helps developers release software faster and safer. The company fosters a collaborative, respectful, and inclusive culture with a focus on team empowerment.
Own and optimize the ingestion layer that scrapes, processes, and lands transaction and listing data from auction houses and marketplaces.
Build monitoring and alerting for pipeline latency, uptime, and source coverage, ensuring a source going quiet pages you immediately.
Modernize storage and processing to cut costs and manual intervention, partnering with pricing, ML, product, and analytics teams.
Alt is unlocking the value of alternative assets, starting with the $5 billion trading card market, by letting collectors buy, sell, vault, and finance their cards in one place. Backed by leaders at Stripe, Coinbase, Seven Seven Six, and pro athletes, we are a remote-first startup with a focus on real-time pricing at scale.
Build core data platform capabilities, developing services and APIs for data access, query execution, ingestion, and synchronization.
Make integrations reusable by building connector frameworks that support different systems without bespoke implementations.
Own execution reliability by building orchestration, retries, checkpointing, and observability for predictable workload recovery.
Tessera Labs is redefining how enterprises adopt and operationalize AI by building multi-agent systems that automate complex business workflows across platforms like SAP and Salesforce. Backed by top venture capital firms and built by leaders from Meta AI and Google Research, they move fast and operate with extreme ownership.
Design, build, and maintain scalable data pipelines and ETL/ELT processes using Databricks and Spark.
Architect and optimize data models and storage solutions for analytics and operational use.
Implement observability, alerting, and data quality monitoring for critical pipelines.
Sonatype provides end-to-end software supply chain security solutions, protecting against malicious open source and managing SBOMs. With over 2,000 organizations and 15 million developers using its platform, it focuses on innovation and security in software development.
Design, build, and own custom data pipelines from source systems into our warehouse.
Build and operate Airflow DAGs for scheduled and event-triggered workflows.
Develop dbt models and a governed semantic layer that analysts, internal tools, and AI agents can trust.
Grüns provides comprehensive nutrition through convenient gummies made from whole-food ingredients. They are a high-growth remote company with a culture focused on ownership, impact, and high standards.
Build and run the data platform, including Snowflake administration, access governance, and ingestion orchestration.
Set technical direction for the medallion architecture, dbt project standards, and reliability practices.
Mentor engineers and analysts, establishing the technical bar for data engineering at Eve.
Eve is redefining legal technology for plaintiff law firms, providing AI-powered solutions that handle cases from intake to resolution. Trusted by over 1000 law firms and backed by $160M from top investors, the company is growing rapidly with a world-class team.
Build the foundational data platform for analytics, risk, and decision-making, migrating siloed data into a unified BigQuery warehouse.
Design and operate reliable ETL/ETL pipelines, consolidate fragmented data sources, and establish standards for ingestion, transformation, and documentation.
Own data quality, observability, and access controls while partnering with Data Science, Engineering, Risk, and Product teams.
Ondo Finance provides institutional-grade, blockchain-enabled investment products and services, developing decentralized finance technology and managing tokenized funds. Founded by Goldman Sachs Digital Assets alumni, the company is backed by top investors like Founders Fund and Coinbase Ventures, is well-capitalized and growing quickly, with a fully remote team across the U.S.
Own core data pipelines end to end, building and operating ELT pipelines that move data from product and third-party systems into Snowflake.
Build the data foundation for AIDE's AI initiatives, making governed data available for agentic AI tooling.
Own data governance and pipeline health, including access controls, data quality, and warehouse cost management.
Benchling is an AI platform for biotech R&D that enables scientists to design experiments, capture structured data, and run AI agents in their workflows. Over 200,000 scientists worldwide trust Benchling, and the company emphasizes AI fluency and innovation in its culture.
Design and ship data infrastructure at the core of modern data platforms, scalable pipelines, and reporting layers.
Work directly with the CEO and a small senior team, owning results from day one.
Build client-facing systems, translating business problems into architecture and delivering results.
Paradox Machines is a data and AI company that helps businesses make their data actually usable. They are a venture-backed company with a small, senior team that values ownership and impact.
Design production data pipelines using Databricks Lakeflow, Autoloader, and Structured Streaming with CI/CD.
Architect Lakehouse solutions including medallion architecture and Unity Catalog for analytics and AI.
Collaborate with product and backend engineers to design data models and APIs that serve the product.
Livefront helps companies design and build world-class digital products. They have worked with major brands and startups, fostering a culture built on respect, mutual trust, and egoless collaboration.
Design and build data pipelines for ML use cases, implementing data versioning and orchestration.
Deploy data scientists' scripts and models from notebooks to production, managing Python dependencies and CI/CD.
Optimize data processing jobs for performance and cost on distributed systems and cloud data services.
NIQ is the world's leading consumer intelligence company, delivering the most complete understanding of consumer buying behavior and revealing new pathways to growth. In 2023, NIQ combined with GfK, bringing together two industry leaders with unparalleled global reach, operating in 100+ markets.
Lead the migration of analytics workflows into BigQuery and build reliable ETL/ELT pipelines.
Consolidate on-chain and off-chain data into a trusted ecosystem and establish data governance standards.
Work closely with Data Science, Backend Engineering, Risk, and Product to translate complex needs into scalable infrastructure.
This company operates in the financial technology sector, building advanced analytics and risk management products using blockchain and crypto data. It is an early-stage, well-capitalized organization with a fully remote team that values autonomy, technical excellence, and strong ownership.
Build and maintain secure connectors across data platforms like Google Drive, Slack, and Claude.
Partner with scientists to integrate lab data collection tools into a unified, queryable platform.
Design pipelines for messy real-world data and own end-to-end infrastructure, from schema to monitoring.
Ohr creates molecules from atoms up using biocatalysis and primordial chemistry, powering rockets and securing industries with cleaner, scalable systems. As an early-stage company with a small, urgent team, it cultivates a culture of vision, hustle, and collaboration.
Design, build, and operate end-to-end batch and incremental ETL/ELT pipelines turning raw data into production-grade APIs, dashboards, and automated alerts.
Develop high-throughput data acquisition from relational databases, APIs, flat files, blockchain nodes, and web text at scale.
Build NLP pipelines for entity extraction, classification, and enrichment feeding downstream risk models.
Inca Digital is a veteran-owned data and intelligence company specializing in digital-asset analytics for exchanges, financial institutions, regulators, and blockchain ecosystems. They operate as a fast-paced, nimble, global, and remote technology company with an asynchronous-first workflow.