Design, build, and maintain data infrastructure and pipelines spanning both batch and real-time workloads.
Build and maintain scalable ETL/ELT pipelines using Python, SQL, Spark, and orchestration frameworks.
Partner with data scientists, ML engineers, and product teams to deliver data products and establish data quality frameworks.
Gemini is a global crypto and Web3 platform founded by Cameron and Tyler Winklevoss in 2014, offering a wide range of crypto products to individuals and institutions in over 70 countries. As a publicly traded company, Gemini is poised to accelerate the vision of reshaping the global financial system with greater scale, reach, and impact.
Lead, coach, and develop a global team of data engineers while guiding technical architecture and delivery.
Build and scale large-scale data pipelines, production datasets, and QA systems using Databricks, Airflow, SQL, and PySpark.
Partner with product and research teams to transform complex data into reliable, production-grade assets for customers and internal applications.
YipitData is a leading market research and analytics firm that analyzes billions of alternative data points to provide actionable insights for the disruptive economy. The company recently raised $475M, operates globally with offices in the US, APAC, and India, and has been recognized by Inc. as a Best Workplace for three consecutive years.
Contribute to the design, development, and operation of core components of Abnormal’s data platform.
Build tools and services that make it easy for other teams to adopt and scale data systems.
Help automate infrastructure and operations to improve reliability, performance, and scalability.
Abnormal protects the humans behind the world's most critical organizations from AI-powered cybercrime. 4,500+ enterprises trust our behavioral AI platform.
Lead an engineering team responsible for delivering reliable, scalable infrastructure supporting the Data Feeds product.
Partner with Product and technical leadership to translate strategic priorities into well-executed engineering roadmaps.
Provide engineering judgment to support the team in making sound architectural and operational decisions.
Chainlink is the industry-standard oracle platform bringing the capital markets onchain and powering the majority of decentralized finance (DeFi). Since inventing decentralized oracle networks, Chainlink has enabled tens of trillions in transaction value and now secures the vast majority of DeFi.
Build, expand, and optimize data infrastructure to create the most accurate dataset of identities and their relationships.
Develop and operate secure, scalable, and reliable data ingestion and ETL/ELT pipelines that meet product requirements.
Design and maintain a data observability framework to ensure data meets strict quality and freshness standards.
SentiLink provides innovative identity and risk solutions, empowering institutions and individuals to transact with confidence. The company is growing rapidly, has verified hundreds of millions of identities, and is backed by top investors like Craft Ventures and Andreessen Horowitz, with offices across the US and India.
Deliver on business-critical outcomes by owning backend and data services end-to-end, from design to production operation.
Design, build, and operate scalable data pipelines and streaming ingestion for high-volume security telemetry.
Contribute to data modeling and warehouse/lakehouse architecture decisions that serve detection, analytics, and product features.
Blackpoint Cyber is the leading provider of world-class cybersecurity threat hunting, detection and remediation technology. Founded by former National Security Agency (NSA) cyber operations experts, the company is in hyper-growth mode, fueled by a recent $190m series C round.
Design and build batch data pipelines that ingest, validate, and transform multi-billion-row datasets.
Model complex real-world data including dimensional models and temporal data.
Develop and operate workloads on lakehouse platforms like Databricks, Spark, and Delta.
Simulmedia builds an advanced TV and streaming advertising platform. They have a team of engineers, data scientists, and designers who are obsessed with building cutting-edge technology.
Design and implement scalable data architectures to meet business needs.
Develop and optimize data pipelines for ingesting and processing large volumes of data.
Provide technical leadership and mentorship to junior engineers.
Oportun is a mission-driven financial services company that provides affordable credit and financial tools. Since inception, it has provided over $21.3 billion in credit and saved members $2.5 billion, and fosters a diverse, inclusive culture.
Design and develop distributed systems for a data platform handling petabytes of data.
Take ownership of components like data ingestion and decoding.
Write code in Kotlin and Go with a focus on good design and performance.
Dune is a collaborative multi-chain analytics platform that makes crypto data accessible. They are a team of ~60 employees working across Europe and eastern US timezones, backed by top investors.
Lead impactful customer technical projects by delivering production-grade systems spanning data engineering, AI, and application development.
Guide strategic customers in implementing end-to-end big data and AI projects, including architecture, design, build, and deployment.
Empower customers by providing architecture guidance and ensuring solutions are secure, scalable, and aligned with Databricks best practices.
Databricks is the data and AI company. More than 10,000 organizations worldwide, including Comcast and Condé Nast, rely on the Databricks Data Intelligence Platform to unify and democratize data, analytics and AI. The company is headquartered in San Francisco and fosters a diverse, inclusive culture.
Own the design, delivery, and operation of major data products including ETL pipelines and storage solutions.
Partner with engineering, product, and business stakeholders to define solutions and drive projects from design to production.
Mentor engineers, lead technical strategy, and participate in on-call rotation for production systems.
Lime is the largest global shared micromobility company, on a mission to make transportation shared, affordable, and carbon-free. It has powered over one billion rides across 30 countries and is a Time 100 Most Influential Company.
Collaborate with product managers, data analysts, and ML engineers to develop pipelines and ETL tasks for insights.
Establish data architecture processes that are scheduled, automated, and replicated as standards.
Manage individual Data Engineers to foster learning, growth, and success.
Doximity is the leading digital platform for U.S. medical professionals, with over 85% of physicians as members. We are a distributed team of doers passionate about improving healthcare.
Build and maintain batch data pipelines ingesting from Workday and legacy systems into a medallion lakehouse on Azure Databricks and Microsoft Fabric.
Develop transformation logic across Bronze/Silver/Gold layers using Spark, Python, and SQL, and implement data-quality checks.
Work within a federated architecture and Unity Catalog security model, contributing to documentation and code reviews.
CampusWorks is a consulting firm that helps higher education institutions leverage technology and managed services to drive transformation and student success. Founded in 1999, it is a large, virtual company with a culture that values work-life balance and employee appreciation.
Design and build scalable backend services and APIs.
Develop integrations with healthcare and financial systems.
Build reliable, observable data ingestion and processing pipelines.
This healthcare AI startup builds products that help providers navigate reimbursement landscapes using machine learning and AI. It has assembled an exceptional technical team and is fast-growing.
Own production data pipelines end-to-end, ensuring reliability, performance, and scalability. - Build and optimize data solutions using PySpark and Foundry tools to process massive datasets. - Champion data quality by designing monitoring and validation frameworks to guarantee trusted data.
We are an IT solutions company providing innovative information and communication technology services. We have grown to over 3,900 employees and are the second largest employer in eastern Slovakia, with a culture focused on continuous improvement and transformation.
Lead the Data Engineering team to design, build, and operate scalable data pipelines powering analytics and AI.
Partner with cross-functional leaders to define technical strategy, roadmap, and execution for the enterprise data platform.
Drive engineering excellence through improved data quality, reliability, observability, and governance.
Omada Health is reverse engineering healthcare delivery in America, focusing on chronic conditions like obesity, diabetes, and hypertension. With over two million members served and a strong remote-first culture, the company has been certified as a Great Place to Work.
You will develop and maintain end-to-end data pipelines and contribute to Samsara's Data Platform for advanced automation and analytics.
You will design, build, and optimize large-scale Spark and PySpark workflows for batch and streaming data processing.
You will build and maintain MCP servers and AI agents, and champion data engineering best practices across the team.
Samsara is the pioneer of the Connected Operations Cloud, enabling organizations to harness IoT data for actionable insights. As a recently public company, they foster a culture of autonomy, support, and rapid career development in a hyper-growth environment.
Build new analyses and support existing ones using SQL and Python.
Apply software engineering principles like version control and continuous integration to the analytics codebase.
Expand our data warehouse with clean data ready for analysis.
Newsela is a leading education technology company dedicated to meaningful classroom learning for every student. It delivers integrated, AI-powered solutions to unlock student engagement and empower teachers.
Own and scale distributed systems for orchestration, APIs, and data paths.
Apply query-engine expertise to optimize work execution and data movement.
Exercise strong technical judgment on ambiguous problems.
Prefect builds and operates resilient, Pythonic orchestration and MCP platforms used for mission-critical workloads. This remote-first company fosters a supportive, high-performance culture that empowers team members to do their best work.