Design, architect, and develop robust GCP data pipelines for healthcare data solutions.
Maintain a holistic view of information assets through documentation and create technical design specs.
Collaborate with cross-functional teams to drive ML and data engineering initiatives that improve patient outcomes.
Egen is a data-first consulting company leveraging Google Cloud and Salesforce to help clients drive impact through data. It is a fast-growing, entrepreneurial firm with a culture focused on learning, innovation, and solving tough problems.
Build and maintain BigQuery data models using Dataform, following medallion architecture patterns (Bronze/Silver/Gold).
Contribute to Looker dashboards and LookML models, working alongside senior engineers and analysts.
Build and maintain robust Python data pipelines with testing, linting, and CI/CD integration.
Kitman Labs is a performance intelligence company that transforms how the sports industry uses data to unlock athlete potential. With over 2000 teams in 50 leagues across 6 continents, the company has assembled a team of top data scientists, sports performance scientists, and product specialists, fostering an innovative and collaborative culture.
Design, build, and maintain scalable data pipelines for ingestion and transformation.
Work with Python, SQL, Apache Airflow, and Google Cloud Platform.
Collaborate with cross-functional teams to deliver reliable and high-quality data solutions.
Our partner is a technology company focused on AI and data solutions. The team values collaboration, innovation, and continuous learning, and offers a remote work environment.
Design and build batch data pipelines that ingest, validate, and transform multi-billion-row datasets.
Model complex real-world data including dimensional models and temporal data.
Develop and operate workloads on lakehouse platforms like Databricks, Spark, and Delta.
Simulmedia builds an advanced TV and streaming advertising platform. They have a team of engineers, data scientists, and designers who are obsessed with building cutting-edge technology.
Support new and existing customers in their data engineering needs, guiding them to make optimal technical decisions.
Build and operationalize complex data solutions, including data governance, security, and quality.
Collaborate across teams to deliver data products and educate end users on analytic environments.
SunnyData is a high-growth consulting company specializing in data engineering and AI, dedicated to the Databricks platform. They foster a collaborative and innovative culture with a focus on customer impact and career growth.
Design, build, and maintain scalable data pipelines and ETL/ELT processes using Databricks and Spark.
Architect and optimize data models and storage solutions for analytics and operational use.
Implement observability, alerting, and data quality monitoring for critical pipelines.
Sonatype provides end-to-end software supply chain security solutions, protecting against malicious open source and managing SBOMs. With over 2,000 organizations and 15 million developers using its platform, it focuses on innovation and security in software development.
Build and maintain secure connectors across data platforms like Google Drive, Slack, and Claude.
Partner with scientists to integrate lab data collection tools into a unified, queryable platform.
Design pipelines for messy real-world data and own end-to-end infrastructure, from schema to monitoring.
Ohr creates molecules from atoms up using biocatalysis and primordial chemistry, powering rockets and securing industries with cleaner, scalable systems. As an early-stage company with a small, urgent team, it cultivates a culture of vision, hustle, and collaboration.
Maintain and monitor Kafka, Hadoop, Presto, and RDBMS systems.
Ingest, validate, and process internal and third-party data flows.
Build Kafka consumers using Spark Streaming for near-real-time aggregation.
PulsePoint sits at the intersection of healthcare and adtech, helping brands interpret health journey signals. We are 300+ employees, growing, and a leading player in the US healthcare ad market.
Develop and maintain scalable, resilient data pipelines using PySpark and distributed processing frameworks.
Modernize legacy processes and design data solutions on AWS, building data products and complex transformations.
Support data quality and optimization initiatives, collaborating with stakeholders to translate requirements into technical solutions.
Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. It focuses on efficiency and fairness in recruitment, promoting diversity and inclusion in the workplace.
Design and implement ETL pipelines using Apache Airflow, BigQuery, Python, and Spark to transform upstream data into curated data assets.
Provide technical leadership and best practices, mentoring other engineers and driving architecture decisions for high-performance systems.
Collaborate cross-functionally with Product Managers and end users to define key business questions and build relevant data sets.
InMarket is a leader in 360-degree consumer intelligence and real-time activation for top brands, offering a data-driven marketing platform. The company has a strong focus on technology and culture, with a commitment to diversity, equity, and inclusion, and offers competitive compensation and benefits.
Design, develop, and maintain scalable ETL/ELT pipelines and data integration processes.
Build and optimize cloud-based data architectures to support analytics and business intelligence initiatives.
Ensure data quality, consistency, governance, and reliability through validation, monitoring, and automated quality checks.
CI&T helps large enterprises transform AI potential into business impact with AI deployment, AI-native execution, and tech-integrated solutions. With 30 years of experience and 8,000 employees across 25 countries, they collaborate to build solutions with real impact.
You will develop and maintain end-to-end data pipelines and contribute to Samsara's Data Platform for advanced automation and analytics.
You will design, build, and optimize large-scale Spark and PySpark workflows for batch and streaming data processing.
You will build and maintain MCP servers and AI agents, and champion data engineering best practices across the team.
Samsara is the pioneer of the Connected Operations Cloud, enabling organizations to harness IoT data for actionable insights. As a recently public company, they foster a culture of autonomy, support, and rapid career development in a hyper-growth environment.
Lead a centralized team of data engineers and applied data scientists to build scalable data infrastructure, machine learning systems, and analytics foundations.
Partner with cross-functional leaders to translate business needs into technical solutions, balancing long-term platform investments with immediate priorities.
Champion data quality, governance, observability, and modern engineering practices across the data ecosystem to drive operational efficiency and AI-powered experiences.
Our partner is a technology company that leverages data and AI to enhance user experiences through scalable data infrastructure and machine learning systems. They foster a collaborative, inclusive, and high-performing culture with a remote-first approach and a focus on professional growth.
Lead the development of the data platform supporting clinical operations, payer reporting, and analytics.
Build and orchestrate ingestion pipelines using Mage, Python, and GCP infrastructure as code.
Ensure security and compliance with PHI, HIPAA, and SOC 2 while collaborating cross-functionally.
Protera Health is a musculoskeletal (MSK) specialty care company that partners with health plans, health systems, and employers to deliver coordinated, measurable MSK care. It is a growing multidisciplinary clinical organization with a technology platform, offering a fully remote, mission-driven culture.
Design and build scalable data pipelines and schemas for client engagements.
Lead data architecture and modeling discussions, weighing tradeoffs.
Mentor teammates and contribute to 8th Light's culture and values.
8th Light is a technology solutions consultancy that partners with organizations to build software, platforms, and agentic products. Founded in 2006, the company fosters an open, collaborative culture grounded in honesty and continuous learning.
Enable efficient data access by creating and maintaining data pipelines.
Collaborate with ML engineers to design and maintain automation for machine learning training, quality assessment, and model release.
Build data infrastructure for analytics, hypothesis testing, and company metrics.
Eneba is building an open, safe, and sustainable marketplace for gamers, supporting close to 20 million active users. We are a growing international team that values data-driven decision making and fosters a healthy data culture.
Support and maintain production GCP data platforms, pipelines, and workflows.
Optimize Airflow DAGs and troubleshoot pipeline failures and performance issues.
Collaborate with cross-functional teams on production issues and enhancements.
Innodata is a global data engineering company that enables the responsible advancement of AI by providing data, evaluation frameworks, and human expertise. With a 36+ year legacy, it delivers high-quality data and outcomes for customers.
Design and develop automated ETL/ELT pipelines to ingest data into Google Cloud Platform.
Implement Medallion Architecture patterns and maintain governed views in BigQuery.
Balance new solution development with production incident resolution and support.
Our partner is a company seeking a senior data professional to design, build, and operate scalable data solutions in a large-scale corporate environment. They collaborate across multiple business areas and combine development with operational support.
Build and maintain batch data pipelines ingesting from Workday and legacy systems into a medallion lakehouse on Azure Databricks and Microsoft Fabric.
Develop transformation logic across Bronze/Silver/Gold layers using Spark, Python, and SQL, and implement data-quality checks.
Work within a federated architecture and Unity Catalog security model, contributing to documentation and code reviews.
CampusWorks is a consulting firm that helps higher education institutions leverage technology and managed services to drive transformation and student success. Founded in 1999, it is a large, virtual company with a culture that values work-life balance and employee appreciation.
Build and maintain data pipelines using Databricks, dbt, Airbyte, and Airflow.
Contribute to data platform architecture and modeling, delivering trusted datasets.
3+ years experience in Data Engineering with strong SQL, Python, and cloud skills.
Ayming is an international consulting firm that supports businesses in digital and technological transformation. The company is expanding its technology, data, and transformation teams in Portugal and fosters a collaborative, innovative work environment.