Prepare clean, documented datasets as a product for AI agents and self-service analytics.
Promova redefines language education for the AI-powered era, making learning accessible and personal worldwide. With over 150 teammates, they blend AI innovation with human creativity to connect people and bridge cultures.
Design, implement, and maintain ETL/ELT pipelines for clean, scalable data flows across multiple systems.
Own data warehouse architecture, including partitioning, access controls, and security in partnership with Security and Infrastructure teams.
Champion data quality and governance, building dashboards and analytics to deliver actionable insights for leadership and cross-functional teams.
CertifID protects life's largest transactions from fraud, helping title companies, law firms, lenders, and consumers safeguard billions of dollars from wire fraud daily. They are a growth-minded team passionate about securing sensitive data and transforming fraud prevention.
Design, develop, and maintain scalable data pipelines that transform complex customer data into actionable business value.
Work across the full data lifecycle, from ingestion and transformation to activation and analytics.
Use modern cloud technologies and AI-assisted development workflows to create reliable, high-performing data solutions.
Jobgether uses an AI-powered matching process to help candidates find roles and connects top-fitting applicants directly with hiring companies. The company fosters a remote, collaborative environment with a focus on efficient recruitment.
Deliver high-quality data sets by curating, consolidating, and manipulating large-scale data sources.
Build high-quality data pipelines and ETL processes on platforms like Snowflake and BigQuery.
Partner with analytics, data science, and machine learning teams to provide data solutions.
Tripadvisor connects people to experiences worth sharing, aiming to be the world's most trusted source for travel. With over 500 million reviews and 390 million monthly visitors, the company is data-driven and offers a unique, global work environment that captures the speed and innovation of a startup.
Build and scale reliable data infrastructure powering analytics, AI-driven decision-making, and smarter growth across n8n.
Own critical parts of the modern data stack, including end-to-end data pipelines using dbt on BigQuery and orchestration with Dagster.
Enable analytics, AI use cases, and marketing attribution by delivering trusted datasets and scalable data products.
n8n is an open workflow orchestration platform built for the new era of AI, giving technical teams the freedom of code with the speed of no-code. The company has grown to a team of over 260 across Europe and the US, backed by top investors, with a culture of builder spirit and remote-first collaboration.
Design, build, and own data pipelines moving data from application and third-party sources into relational databases.
Build and ship production AI systems, including data infrastructure and evaluation systems for feature reliability.
Set standards for data work through clear writing, early context sharing, and team efficiency.
Dscout builds a flexible UX research platform trusted by top brands in finance, healthcare, and tech. They are a remote-first team of passionate professionals that prioritize learning, diversity, and inclusion.
Build the data platform by designing and maintaining scalable ELT/ETL pipelines and owning the cloud data warehouse.
Ensure data quality and governance through clean transformation layers, testing, monitoring, and security practices.
Partner across engineering and product teams to support product analytics, experimentation, and business reporting.
An international product company building SaaS products, AI-powered solutions, and modern web and mobile applications for global markets. They are scaling and investing in a modern data platform, with a collaborative, product-focused team.
Design, build, and deploy scalable ETL/ELT pipelines from diverse source systems into Snowflake Data Cloud.
Manage and optimize data flows within an AWS environment, leveraging Databricks and Python for high-scale processing.
Implement a Semantic Layer via dbt to standardize business logic, metrics, and dimensions for all downstream consumers.
Instructure creates intuitive products that simplify learning and personal development, amplifying the power of people to grow and succeed. They offer competitive benefits and a culture rooted in inclusivity, support, and meaningful connection.
Design, develop, and maintain scalable ETL/ELT pipelines for enterprise data integration and analytics.
Optimize data platforms like Databricks Unity Catalog and SQL Server Managed Instances, implementing Lakehouse architecture.
Apply data quality controls, governance standards, and collaborate with cross-functional teams to deliver high-quality data solutions.
Ardent supports the federal government in national security and defense priorities, helping protect the nation and advance critical technologies. The company offers competitive pay, comprehensive health coverage, and a culture that values dedication and adaptability, employing purpose-driven innovators and veterans.
Build and maintain BigQuery data models using Dataform, following medallion architecture patterns.
Build robust Python data pipelines with testing, linting, and CI/CD integration.
Support Data Scientists in moving work from notebook to production pipeline.
Kitman Labs is a performance intelligence company disrupting the sports industry by using data to unlock athlete potential. Their team of top data scientists, performance scientists, and engineers serves over 2000 teams in 50 leagues globally, fostering an innovative and world-class culture.
Collaborate with Data Science, Product Managers, and Software Engineers to build robust ETL pipelines for user-facing features.
Contribute to architecture decisions, observability tooling, and data quality initiatives to keep the platform robust.
Enforce engineering best practices across the AI/ML org, including code quality, testing, and documentation.
Federato is an AI-native platform for insurance, enabling insurers to provide affordable coverage for climate, cyber, and social inflation risks. It is a small, well-funded company backed by the investors behind Salesforce, Veeva, and Zoom, with a culture focused on first principles, learning, and fun.
Develop and implement Snowflake-based governance solutions for data quality, security, and compliance.
Build automated privacy processes and support data cleanup initiatives to maintain integrity.
Collaborate with cross-functional teams to design scalable data architecture and pipeline solutions.
Blend is a premier AI services provider, co-creating meaningful impact through data science, AI, and technology. We harness world-class people and data-driven strategy to unlock value for clients, with a focus on innovation and fulfilling work.
Build and maintain ETL/ELT data pipelines for large-scale organizations.
Develop data solutions using SQL, Python, and tools such as Snowflake and AWS.
Collaborate with non-technical stakeholders to translate business requirements into scalable technical solutions.
We are a trusted technology and transformation solutions provider for global brands, with over 30 years of experience. We have more than 3,300 employees, €350/£300m in revenue, and a values-driven culture that has earned us recognition as a Great Place to Work.
Design and build scalable data pipelines, clean room environments, and privacy-safe integrations for NBCUniversal’s data collaboration ecosystem.
Implement identity resolution logic and configure secure, role-based access controls across data platforms.
Optimize query performance and operational reliability, including monitoring, cost tracking, and incident response.
NBCUniversal is one of the world's leading media and entertainment companies, creating and distributing content across film, television, and streaming. A subsidiary of Comcast Corporation, it champions an inclusive culture and has a rich tradition of community service.
Design, build, and maintain robust, scalable ELT/ETL data pipelines from various source systems into cloud data platforms.
Perform data modeling, including dimensional modeling, and build transformation layers using dbt to create analytics-ready datasets.
Support operational reliability, monitor data pipelines, and ensure SLAs for timeliness, freshness, and accuracy.
Troveo builds the data platform that AI labs and model builders need to train the next generation of models. Backed by top investors, we’re a small, high-impact team solving one of the biggest bottlenecks in AI development.
Design and optimize large-scale ETL pipelines using Python, PySpark, SQL, DBT, and cloud-based data platforms.
Define technical vision and architecture for data integration solutions, ensuring scalability and reliability.
Lead technical initiatives, mentor engineers, and collaborate with cross-functional teams to deliver high-quality data solutions.
Jobgether uses AI-powered matching to connect candidates with job opportunities at partner companies. They focus on efficient, objective recruitment processes for a fast-growing remote-first environment.
Design, build, and maintain scalable data pipelines using Python, SQL, Snowflake, Dagster, dbt, and AWS.
Own end-to-end data engineering projects from ingestion through to analytics enablement.
Improve monitoring, alerting, and data quality across key pipelines.
Midnite is a next-generation sports betting and gaming platform built for a new wave of players. Over 400,000 players have joined, and the team operates with high ownership and fast iteration in a scale-up environment.
Lead the design and implementation of the enterprise data warehouse on Databricks, migrating from BigQuery.
Build scalable ETL/ELT pipelines and dimensional data models to support internal analytics.
Establish data warehouse standards, governance, and best practices across the organization.
Bloomerang is a nonprofit giving platform that helps tens of thousands of nonprofits raise more, recruit more, and retain more. The company has a mission-driven culture built on core values of Simplify, Care, and Act.
Design, build, and maintain scalable batch and streaming data pipelines.
Develop reliable ETL/ELT workflows using Python, Spark, and modern orchestration tools.
Improve data quality, validation, monitoring, and observability across the platform.
The company is a Berlin-based, remote-first technology company building advanced market intelligence and software solutions for the automotive industry. It operates in a stable growth phase with an established product and a strong technical team.
Design, build, and evolve Data Warehouses, Data Marts, and enterprise data models to support analytics and business decision-making.
Design, implement, and optimize scalable ELT pipelines and data processing solutions, ensuring reliability, performance, and maintainability.
Drive technical decisions related to data architecture, platform evolution, data modeling approaches, and integration patterns.
Sigma Software is building a scalable, reliable, and secure corporate Data Platform to power analytics and decision-making across the company. The company supports professional growth, knowledge sharing, and modern tools for high-quality solutions.