Design, develop, and maintain scalable ETL/ELT pipelines and data integration processes.
Build and optimize cloud-based data architectures to support analytics and business intelligence initiatives.
Ensure data quality, consistency, governance, and reliability through validation, monitoring, and automated quality checks.
CI&T helps large enterprises transform AI potential into business impact with AI deployment, AI-native execution, and tech-integrated solutions. With 30 years of experience and 8,000 employees across 25 countries, they collaborate to build solutions with real impact.
Deliver high-quality data sets by curating, consolidating, and manipulating large-scale data sources.
Build high-quality data pipelines and ETL processes on platforms like Snowflake and BigQuery.
Partner with analytics, data science, and machine learning teams to provide data solutions.
Tripadvisor connects people to experiences worth sharing, aiming to be the world's most trusted source for travel. With over 500 million reviews and 390 million monthly visitors, the company is data-driven and offers a unique, global work environment that captures the speed and innovation of a startup.
Own the Airflow codebase end-to-end, building reusable templates and enforcing standards.
Build and extend large-scale Spark pipelines on AWS Glue and support migration to Databricks.
Drive data model improvements around commercial pharma data with focus on structure and lineage.
Veeva Systems is a mission-driven organization and pioneer in industry cloud, helping life sciences companies bring therapies to patients faster. As one of the fastest-growing SaaS companies in history, it surpassed $3B in revenue and values doing the right thing, customer success, employee success, and speed.
Own data pipelines end to end - ingestion, transformation, warehousing, and delivery into BI.
Build and maintain dbt models with tests, documentation, and clear contracts.
Own product analytics: instrumentation, event tracking, usage models, and metrics for GTM and product teams.
Cortex is the Engineering Operations Platform providing visibility, governance, and golden paths for engineering organizations. We are a team of 80 passionate individuals excited about building a product that developers love.
You will develop and maintain end-to-end data pipelines and contribute to Samsara's Data Platform for advanced automation and analytics.
You will design, build, and optimize large-scale Spark and PySpark workflows for batch and streaming data processing.
You will build and maintain MCP servers and AI agents, and champion data engineering best practices across the team.
Samsara is the pioneer of the Connected Operations Cloud, enabling organizations to harness IoT data for actionable insights. As a recently public company, they foster a culture of autonomy, support, and rapid career development in a hyper-growth environment.
Design and build batch data pipelines that ingest, validate, and transform multi-billion-row datasets.
Model complex real-world data including dimensional models and temporal data.
Develop and operate workloads on lakehouse platforms like Databricks, Spark, and Delta.
Simulmedia builds an advanced TV and streaming advertising platform. They have a team of engineers, data scientists, and designers who are obsessed with building cutting-edge technology.
Design and implement scalable lakehouse architectures using Bronze, Silver, and Gold patterns with technologies like Kafka, Spark, and Airflow.
Build and manage high-performance streaming and batch data pipelines and open table format solutions such as Iceberg, Delta Lake, and Hudi.
Optimize distributed query environments, maintain platform reliability, and partner with data scientists and product teams to deliver impactful data capabilities.
The partner company is building scalable data platforms that power analytics, machine learning, and next-generation products. They offer a collaborative and supportive work environment focused on innovation and continuous learning.
Design and implement robust, scalable data ingestion and transformation pipelines using Databricks, PySpark, and distributed processing.
Implement Delta Lake principles focusing on CDC and schema evolution, and integrate data quality frameworks within CI/CD pipelines.
Develop and optimize complex SQL and Python scripts, handling diverse data sources and supporting data governance solutions.
Mobile Wave Solutions is a professional services company specializing in software development as a service. With a team of over 120 engineers, we deliver scalable, high-quality software that empowers our global clients to innovate and grow.
Lead the design and implementation of the enterprise data warehouse on Databricks, migrating from BigQuery.
Build scalable ETL/ELT pipelines and dimensional data models to support internal analytics.
Establish data warehouse standards, governance, and best practices across the organization.
Bloomerang is a nonprofit giving platform that helps tens of thousands of nonprofits raise more, recruit more, and retain more. The company has a mission-driven culture built on core values of Simplify, Care, and Act.
Design, build, and maintain robust, scalable ELT/ETL data pipelines from various source systems into cloud data platforms.
Perform data modeling, including dimensional modeling, and build transformation layers using dbt to create analytics-ready datasets.
Support operational reliability, monitor data pipelines, and ensure SLAs for timeliness, freshness, and accuracy.
Troveo builds the data platform that AI labs and model builders need to train the next generation of models. Backed by top investors, we’re a small, high-impact team solving one of the biggest bottlenecks in AI development.
Translate business goals into concrete data solutions.
Collaborate with peers, manage multiple priorities, and communicate project status.
Apply best practices in data pipelines, schemas, and system design.
DEPT is a growth invention company helping ambitious brands grow at the intersection of technology and marketing. With over 4,000 specialists, it values collaboration, curiosity, and getting things done.
Lead, coach, and develop a global team of data engineers while guiding technical architecture and delivery.
Build and scale large-scale data pipelines, production datasets, and QA systems using Databricks, Airflow, SQL, and PySpark.
Partner with product and research teams to transform complex data into reliable, production-grade assets for customers and internal applications.
YipitData is a leading market research and analytics firm that analyzes billions of alternative data points to provide actionable insights for the disruptive economy. The company recently raised $475M, operates globally with offices in the US, APAC, and India, and has been recognized by Inc. as a Best Workplace for three consecutive years.
Build and maintain batch data pipelines ingesting from Workday and legacy systems into a medallion lakehouse on Azure Databricks and Microsoft Fabric.
Develop transformation logic across Bronze/Silver/Gold layers using Spark, Python, and SQL, and implement data-quality checks.
Work within a federated architecture and Unity Catalog security model, contributing to documentation and code reviews.
CampusWorks is a consulting firm that helps higher education institutions leverage technology and managed services to drive transformation and student success. Founded in 1999, it is a large, virtual company with a culture that values work-life balance and employee appreciation.
Design, build, and maintain scalable data pipelines for ingestion and transformation.
Work with Python, SQL, Apache Airflow, and Google Cloud Platform.
Collaborate with cross-functional teams to deliver reliable and high-quality data solutions.
Our partner is a technology company focused on AI and data solutions. The team values collaboration, innovation, and continuous learning, and offers a remote work environment.
Lead the modernization of the data platform by migrating legacy Hadoop, Spark, and Impala pipelines to a scalable Databricks architecture.
Accelerate migration efforts using AI coding assistants like Claude and Codex to convert SQL and modernize ETL workflows.
Design and optimize scalable data pipelines with Airflow and Databricks, ensuring reliability, cost-efficiency, and data quality.
LivePerson is a leader in enterprise conversational AI and digital transformation, powering nearly a billion conversational interactions monthly. Fast Company named them the #1 Most Innovative AI Company, and they foster a remote-first, innovative culture.
Collaborate with Data Science, Product Managers, and Software Engineers to build robust ETL pipelines for user-facing features.
Contribute to architecture decisions, observability tooling, and data quality initiatives to keep the platform robust.
Enforce engineering best practices across the AI/ML org, including code quality, testing, and documentation.
Federato is an AI-native platform for insurance, enabling insurers to provide affordable coverage for climate, cyber, and social inflation risks. It is a small, well-funded company backed by the investors behind Salesforce, Veeva, and Zoom, with a culture focused on first principles, learning, and fun.
Design, build, and maintain data infrastructure and pipelines spanning both batch and real-time workloads.
Build and maintain scalable ETL/ELT pipelines using Python, SQL, Spark, and orchestration frameworks.
Partner with data scientists, ML engineers, and product teams to deliver data products and establish data quality frameworks.
Gemini is a global crypto and Web3 platform founded by Cameron and Tyler Winklevoss in 2014, offering a wide range of crypto products to individuals and institutions in over 70 countries. As a publicly traded company, Gemini is poised to accelerate the vision of reshaping the global financial system with greater scale, reach, and impact.
Design, build, and maintain scalable data pipelines and systems using Python, PySpark, and AWS-native services.
Develop and optimize data platforms on AWS, leveraging technologies like Redshift, Glue, and Athena.
Mentor junior engineers and collaborate with cross-functional teams in an Agile environment.
Capital Technology Group provides expert consulting services in software development, digital transformation, and data analytics. Recognized as a Top Workplace by The Washington Post, they foster a culture where employees feel trusted and empowered.
Build, expand, and optimize data infrastructure to create the most accurate dataset of identities and their relationships.
Develop and operate secure, scalable, and reliable data ingestion and ETL/ELT pipelines that meet product requirements.
Design and maintain a data observability framework to ensure data meets strict quality and freshness standards.
SentiLink provides innovative identity and risk solutions, empowering institutions and individuals to transact with confidence. The company is growing rapidly, has verified hundreds of millions of identities, and is backed by top investors like Craft Ventures and Andreessen Horowitz, with offices across the US and India.