Design and migrate scalable data solutions using modern cloud technologies like AWS and Data Mesh.
Convert Alteryx workflows to AWS-native architectures such as AWS Glue and improve performance.
Build and maintain reliable data pipelines using distributed processing tools like Spark and Airflow.
The partner company provides data engineering and cloud solutions aimed at helping organizations modernize their data platforms. They are a collaborative team of technology professionals focused on innovation and scalable architectures.
Design and build scalable batch data pipelines using AWS Glue, Lambda, and Step Functions.
Optimize analytical data models and query layers with Amazon Athena, Redshift, and ClickHouse.
Collaborate with backend engineers to integrate data workflows with microservices and event-driven architectures.
The company is a technology organization focused on building scalable data platforms and analytics solutions. They operate with a distributed international team and emphasize modern engineering practices.
Design, develop, and maintain scalable ETL/ELT pipelines and data integration processes.
Build and optimize cloud-based data architectures to support analytics and business intelligence initiatives.
Ensure data quality, consistency, governance, and reliability through validation, monitoring, and automated quality checks.
CI&T helps large enterprises transform AI potential into business impact with AI deployment, AI-native execution, and tech-integrated solutions. With 30 years of experience and 8,000 employees across 25 countries, they collaborate to build solutions with real impact.
Lead a team of data engineers to design and optimize scalable data solutions.
Architect and implement large-scale streaming and batch data processing using Kafka, Spark, and Hadoop.
Collaborate with cross-functional teams to define data strategies and ensure data governance and quality.
The company is a technology-focused organization that builds scalable data platforms. It offers a fully remote work environment with a collaborative culture and international teams.
Design, develop, and maintain scalable ETL/ELT pipelines
Build and optimize data processing solutions using Python and Apache Spark
Develop and support real-time data streaming applications using Kafka
Sigma Software develops modern digital solutions for global businesses. They foster an international remote-first environment where engineers can grow, innovate, and influence technical decisions.
Design, build, and maintain scalable data pipelines and systems using Python, PySpark, and AWS-native services.
Develop and optimize data platforms on AWS, leveraging technologies like Redshift, Glue, and Athena.
Mentor junior engineers and collaborate with cross-functional teams in an Agile environment.
Capital Technology Group provides expert consulting services in software development, digital transformation, and data analytics. Recognized as a Top Workplace by The Washington Post, they foster a culture where employees feel trusted and empowered.
Design, build, and maintain scalable data pipelines for ingestion and transformation.
Work with Python, SQL, Apache Airflow, and Google Cloud Platform.
Collaborate with cross-functional teams to deliver reliable and high-quality data solutions.
Our partner is a technology company focused on AI and data solutions. The team values collaboration, innovation, and continuous learning, and offers a remote work environment.
Design and deliver near real-time data solutions for the analytics platform.
Analyze business needs, optimize data models, and identify slow queries for performance improvement.
Write clean, scalable code using Scala, Python, and SQL while mentoring team members.
Aircall is an AI-powered customer communications platform used by 22,000+ companies worldwide. It is a unicorn startup with offices across multiple countries and a focus on innovation and collaboration.
Design and implement scalable lakehouse architectures using Bronze, Silver, and Gold patterns with technologies like Kafka, Spark, and Airflow.
Build and manage high-performance streaming and batch data pipelines and open table format solutions such as Iceberg, Delta Lake, and Hudi.
Optimize distributed query environments, maintain platform reliability, and partner with data scientists and product teams to deliver impactful data capabilities.
The partner company is building scalable data platforms that power analytics, machine learning, and next-generation products. They offer a collaborative and supportive work environment focused on innovation and continuous learning.
You will develop and maintain end-to-end data pipelines and contribute to Samsara's Data Platform for advanced automation and analytics.
You will design, build, and optimize large-scale Spark and PySpark workflows for batch and streaming data processing.
You will build and maintain MCP servers and AI agents, and champion data engineering best practices across the team.
Samsara is the pioneer of the Connected Operations Cloud, enabling organizations to harness IoT data for actionable insights. As a recently public company, they foster a culture of autonomy, support, and rapid career development in a hyper-growth environment.
Build and maintain batch data pipelines ingesting from Workday and legacy systems into a medallion lakehouse on Azure Databricks and Microsoft Fabric.
Develop transformation logic across Bronze/Silver/Gold layers using Spark, Python, and SQL, and implement data-quality checks.
Work within a federated architecture and Unity Catalog security model, contributing to documentation and code reviews.
CampusWorks is a consulting firm that helps higher education institutions leverage technology and managed services to drive transformation and student success. Founded in 1999, it is a large, virtual company with a culture that values work-life balance and employee appreciation.
Own the design, delivery, and operation of major data products including ETL pipelines and storage solutions.
Partner with engineering, product, and business stakeholders to define solutions and drive projects from design to production.
Mentor engineers, lead technical strategy, and participate in on-call rotation for production systems.
Lime is the largest global shared micromobility company, on a mission to make transportation shared, affordable, and carbon-free. It has powered over one billion rides across 30 countries and is a Time 100 Most Influential Company.
Design and implement scalable data architectures to meet business needs.
Develop and optimize data pipelines for ingesting and processing large volumes of data.
Provide technical leadership and mentorship to junior engineers.
Oportun is a mission-driven financial services company that provides affordable credit and financial tools. Since inception, it has provided over $21.3 billion in credit and saved members $2.5 billion, and fosters a diverse, inclusive culture.
Design and maintain scalable data pipelines using Microsoft Fabric, Databricks, and Azure services.
Optimize ETL/ELT processes for performance and cost efficiency.
Collaborate with cross-functional teams to transform data requirements into robust solutions.
The company builds modern cloud-based data platforms to power enterprise analytics and decision-making. It offers a fully remote, collaborative environment with global stakeholder interaction and opportunities for professional growth.
Design and implement robust, scalable data ingestion and transformation pipelines using Databricks, PySpark, and distributed processing.
Implement Delta Lake principles focusing on CDC and schema evolution, and integrate data quality frameworks within CI/CD pipelines.
Develop and optimize complex SQL and Python scripts, handling diverse data sources and supporting data governance solutions.
Mobile Wave Solutions is a professional services company specializing in software development as a service. With a team of over 120 engineers, we deliver scalable, high-quality software that empowers our global clients to innovate and grow.
Design and build batch data pipelines that ingest, validate, and transform multi-billion-row datasets.
Model complex real-world data including dimensional models and temporal data.
Develop and operate workloads on lakehouse platforms like Databricks, Spark, and Delta.
Simulmedia builds an advanced TV and streaming advertising platform. They have a team of engineers, data scientists, and designers who are obsessed with building cutting-edge technology.
Build, maintain, and run CI/CD pipelines and infrastructure-as-code for the platform and services.
Provision and manage cloud-based Spark clusters and distributed data processing environments.
Investigate and resolve data pipeline issues while monitoring systems and managing cloud costs.
Smile Digital Health provides a FHIR-based data liberation platform for healthcare stakeholders to collect and exchange data. They were recognized as #19 on Deloitte's Technology Fast 50 Ranking for 2024 and operate in over 20 countries, fostering a culture of respect, inclusion, and diversity.
Support Snowflake-to-Databricks data migration activities.
Design and implement scalable data pipelines using Snowflake, Databricks, SQL, Python, and dbt.
Optimize data workflows for performance, scalability, cost efficiency, and maintainability.
Hiflylabs is a Budapest-based company delivering Data Warehouses, Business Intelligence, and Data Analytics solutions. With over 250 employees and 10 years of experience, they serve financial, telecommunication, and energy sectors, fostering a culture of innovation and collaboration.
Design and build batch-oriented data pipelines and ETL workflows using AWS Glue, Lambda, and Step Functions.
Develop and optimize ingestion pipelines and analytical data models with Athena, Redshift, and ClickHouse.
Collaborate with backend engineers to integrate data workflows with microservices and ensure data quality and observability.
SavvyMoney is a US-based financial technology company providing integrated credit score and personal finance solutions to over 1,600 bank and credit union partners. It was recognized as a top workplace and Inc. 5000 fastest growing company, with a distributed team across the US, Canada, Europe, and India.
Own the Airflow codebase end-to-end, building reusable templates and enforcing standards.
Build and extend large-scale Spark pipelines on AWS Glue and support migration to Databricks.
Drive data model improvements around commercial pharma data with focus on structure and lineage.
Veeva Systems is a mission-driven organization and pioneer in industry cloud, helping life sciences companies bring therapies to patients faster. As one of the fastest-growing SaaS companies in history, it surpassed $3B in revenue and values doing the right thing, customer success, employee success, and speed.