Design and build batch data pipelines that ingest, validate, and transform multi-billion-row datasets.
Model complex real-world data including dimensional models and temporal data.
Develop and operate workloads on lakehouse platforms like Databricks, Spark, and Delta.
Simulmedia builds an advanced TV and streaming advertising platform. They have a team of engineers, data scientists, and designers who are obsessed with building cutting-edge technology.
Lead, coach, and develop a global team of data engineers while guiding technical architecture and delivery.
Build and scale large-scale data pipelines, production datasets, and QA systems using Databricks, Airflow, SQL, and PySpark.
Partner with product and research teams to transform complex data into reliable, production-grade assets for customers and internal applications.
YipitData is a leading market research and analytics firm that analyzes billions of alternative data points to provide actionable insights for the disruptive economy. The company recently raised $475M, operates globally with offices in the US, APAC, and India, and has been recognized by Inc. as a Best Workplace for three consecutive years.
Build and maintain batch data pipelines ingesting from Workday and legacy systems into a medallion lakehouse on Azure Databricks and Microsoft Fabric.
Develop transformation logic across Bronze/Silver/Gold layers using Spark, Python, and SQL, and implement data-quality checks.
Work within a federated architecture and Unity Catalog security model, contributing to documentation and code reviews.
CampusWorks is a consulting firm that helps higher education institutions leverage technology and managed services to drive transformation and student success. Founded in 1999, it is a large, virtual company with a culture that values work-life balance and employee appreciation.
Design and maintain scalable batch and near real-time data pipelines in Databricks, integrating data from various sources.
Build unified customer profiles and support identity resolution, deduplication, enrichment, and Golden Record creation.
Enable CRM and marketing teams to access trusted, activation-ready customer data with minimal latency.
Massive Rocket is a rapidly scaling Braze and Snowflake agency that transforms how digital marketing, product, and engineering teams connect. They have grown quickly in five years and are aiming for $100M in revenue, with a culture of ownership, collaboration, and growth.
Design, build, and maintain data infrastructure and pipelines spanning both batch and real-time workloads.
Build and maintain scalable ETL/ELT pipelines using Python, SQL, Spark, and orchestration frameworks.
Partner with data scientists, ML engineers, and product teams to deliver data products and establish data quality frameworks.
Gemini is a global crypto and Web3 platform founded by Cameron and Tyler Winklevoss in 2014, offering a wide range of crypto products to individuals and institutions in over 70 countries. As a publicly traded company, Gemini is poised to accelerate the vision of reshaping the global financial system with greater scale, reach, and impact.
Design and build ingestion pipelines from enterprise source systems into the Databricks lakehouse using Delta Lake, owning Bronze-layer ingestion and building Silver-layer pipelines for cleansing and standardization.
Implement and maintain Databricks platform constructs for secure delivery, including catalogs, schemas, service principals, and job orchestration, and build CI/CD pipelines for data platform assets.
Apply data classification and segregation requirements within pipeline design, build data quality controls reflecting business meaning, and partner with teams to ensure reliable, governed data for downstream use.
Shield AI is a venture-backed defense-tech company founded in 2015 with the mission of protecting service members and civilians with intelligent systems, developing products like Hivemind autonomy software and V-BAT and X-BAT aircraft. The company has offices and facilities across the U.S., Europe, the Middle East, and Asia-Pacific, and its technology actively supports operations worldwide.
Design, develop, and maintain scalable, production-ready data pipelines and data products using Spark (Python/SQL) in a Databricks environment.
Lead the integration and transformation of complex data from diverse DoD and federal health systems into reliable, reusable data products.
Provide technical guidance and mentorship to other engineers, helping teams navigate complex technical challenges.
540 is a forward-thinking company that delivers innovative technology solutions for government missions. The team has a culture of breaking down barriers and solving mission-critical problems.
Design, build, and operate scalable data pipelines using modern ETL/ELT approaches on cloud-native platforms.
Work with technologies such as Databricks, PostgreSQL, and distributed processing frameworks to deliver production-ready solutions.
Collaborate with analysts, data scientists, and AI engineers while mentoring junior team members.
Version 1 is a technology and transformation solutions company trusted by global brands. With over 3,300 employees and €350m revenue, they are a values-driven employer focused on employee wellbeing and professional growth.
Design and implement scalable data architectures to meet business needs.
Develop and optimize data pipelines for ingesting and processing large volumes of data.
Provide technical leadership and mentorship to junior engineers.
Oportun is a mission-driven financial services company that provides affordable credit and financial tools. Since inception, it has provided over $21.3 billion in credit and saved members $2.5 billion, and fosters a diverse, inclusive culture.
Design, develop, and maintain scalable ETL/ELT pipelines and data integration processes.
Build and optimize cloud-based data architectures to support analytics and business intelligence initiatives.
Ensure data quality, consistency, governance, and reliability through validation, monitoring, and automated quality checks.
CI&T helps large enterprises transform AI potential into business impact with AI deployment, AI-native execution, and tech-integrated solutions. With 30 years of experience and 8,000 employees across 25 countries, they collaborate to build solutions with real impact.
You will develop and maintain end-to-end data pipelines and contribute to Samsara's Data Platform for advanced automation and analytics.
You will design, build, and optimize large-scale Spark and PySpark workflows for batch and streaming data processing.
You will build and maintain MCP servers and AI agents, and champion data engineering best practices across the team.
Samsara is the pioneer of the Connected Operations Cloud, enabling organizations to harness IoT data for actionable insights. As a recently public company, they foster a culture of autonomy, support, and rapid career development in a hyper-growth environment.
Own the design, delivery, and operation of major data products including ETL pipelines and storage solutions.
Partner with engineering, product, and business stakeholders to define solutions and drive projects from design to production.
Mentor engineers, lead technical strategy, and participate in on-call rotation for production systems.
Lime is the largest global shared micromobility company, on a mission to make transportation shared, affordable, and carbon-free. It has powered over one billion rides across 30 countries and is a Time 100 Most Influential Company.
Build, expand, and optimize data infrastructure to create the most accurate dataset of identities and their relationships.
Develop and operate secure, scalable, and reliable data ingestion and ETL/ELT pipelines that meet product requirements.
Design and maintain a data observability framework to ensure data meets strict quality and freshness standards.
SentiLink provides innovative identity and risk solutions, empowering institutions and individuals to transact with confidence. The company is growing rapidly, has verified hundreds of millions of identities, and is backed by top investors like Craft Ventures and Andreessen Horowitz, with offices across the US and India.
Design, build, and maintain scalable data pipelines and platform capabilities to power analytics, AI/ML, and healthcare products.
Implement data quality, governance, and security controls, ensuring healthcare data is handled securely and in compliance.
Provide technical leadership, mentor engineers, and collaborate across teams to translate requirements into scalable data solutions.
Experity is a mission-driven team transforming on-demand healthcare across the U.S., empowering urgent care clinics with industry-leading software. They foster a culture of care, growth, and celebration with day-one benefits, career development, and a supportive team environment.
Own the Airflow codebase end-to-end, building reusable templates and enforcing standards.
Build and extend large-scale Spark pipelines on AWS Glue and support migration to Databricks.
Drive data model improvements around commercial pharma data with focus on structure and lineage.
Veeva Systems is a mission-driven organization and pioneer in industry cloud, helping life sciences companies bring therapies to patients faster. As one of the fastest-growing SaaS companies in history, it surpassed $3B in revenue and values doing the right thing, customer success, employee success, and speed.
Deliver high-quality data sets by curating, consolidating, and manipulating large-scale data sources.
Build high-quality data pipelines and ETL processes on platforms like Snowflake and BigQuery.
Partner with analytics, data science, and machine learning teams to provide data solutions.
Tripadvisor connects people to experiences worth sharing, aiming to be the world's most trusted source for travel. With over 500 million reviews and 390 million monthly visitors, the company is data-driven and offers a unique, global work environment that captures the speed and innovation of a startup.
Build, maintain, and optimize scalable data infrastructure for AI/ML features.
Partner with ML Engineers, Data Scientists, and Software Engineers to transform healthcare data into high-quality datasets.
Contribute to data pipelines, improve data quality, and ensure reliable, performant, and well-governed data for AI systems.
Tebra is the only all-in-one EHR+ platform built exclusively for independent healthcare practices. More than 42,000 private practices trust Tebra to streamline operations, increase revenue, and reduce burnout, with a culture that values starting with the customer, keeping it simple, staying entrepreneurial, being better together, and celebrating success.
Design, build, and maintain scalable data pipelines and systems using Python, PySpark, and AWS-native services.
Develop and optimize data platforms on AWS, leveraging technologies like Redshift, Glue, and Athena.
Mentor junior engineers and collaborate with cross-functional teams in an Agile environment.
Capital Technology Group provides expert consulting services in software development, digital transformation, and data analytics. Recognized as a Top Workplace by The Washington Post, they foster a culture where employees feel trusted and empowered.
Collaborate with Data Science, Product Managers, and Software Engineers to build robust ETL pipelines for user-facing features.
Contribute to architecture decisions, observability tooling, and data quality initiatives to keep the platform robust.
Enforce engineering best practices across the AI/ML org, including code quality, testing, and documentation.
Federato is an AI-native platform for insurance, enabling insurers to provide affordable coverage for climate, cyber, and social inflation risks. It is a small, well-funded company backed by the investors behind Salesforce, Veeva, and Zoom, with a culture focused on first principles, learning, and fun.
Design and build scalable data pipelines and schemas for client engagements.
Lead data architecture and modeling discussions, weighing tradeoffs.
Mentor teammates and contribute to 8th Light's culture and values.
8th Light is a technology solutions consultancy that partners with organizations to build software, platforms, and agentic products. Founded in 2006, the company fosters an open, collaborative culture grounded in honesty and continuous learning.