Build the foundational data platform for analytics, risk, and decision-making, migrating siloed data into a unified BigQuery warehouse.
Design and operate reliable ETL/ETL pipelines, consolidate fragmented data sources, and establish standards for ingestion, transformation, and documentation.
Own data quality, observability, and access controls while partnering with Data Science, Engineering, Risk, and Product teams.
Own data pipelines end to end - ingestion, transformation, warehousing, and delivery into BI.
Build and maintain dbt models with tests, documentation, and clear contracts.
Own product analytics: instrumentation, event tracking, usage models, and metrics for GTM and product teams.
Cortex is the Engineering Operations Platform providing visibility, governance, and golden paths for engineering organizations. We are a team of 80 passionate individuals excited about building a product that developers love.
Design, build, and maintain data infrastructure and pipelines spanning both batch and real-time workloads.
Build and maintain scalable ETL/ELT pipelines using Python, SQL, Spark, and orchestration frameworks.
Partner with data scientists, ML engineers, and product teams to deliver data products and establish data quality frameworks.
Gemini is a global crypto and Web3 platform founded by Cameron and Tyler Winklevoss in 2014, offering a wide range of crypto products to individuals and institutions in over 70 countries. As a publicly traded company, Gemini is poised to accelerate the vision of reshaping the global financial system with greater scale, reach, and impact.
Design and develop automated ETL/ELT pipelines to ingest data into Google Cloud Platform.
Implement Medallion Architecture patterns and maintain governed views in BigQuery.
Balance new solution development with production incident resolution and support.
Our partner is a company seeking a senior data professional to design, build, and operate scalable data solutions in a large-scale corporate environment. They collaborate across multiple business areas and combine development with operational support.
Own external market data ingestion end to end, from acquisition to warehouse delivery.
Build entity resolution and canonical item models to map messy inbound records.
Set schema contracts and maintain dbt models across marketplace, vault, payments, and event data.
We are a collectibles trading platform that enables instant liquidity for items like trading cards and comics, with all assets securely vaulted and insured. We are a remote-first startup focused on growth and innovation, hiring across all functions.
Design, develop, and maintain Google BigQuery data warehouse solutions.
Implement data pipelines and ETL processes to support data integration and analytics.
Collaborate with cross-functional teams to gather requirements and deliver data solutions.
BarkleyOKRP is a creative agency specializing in advertising and marketing, committed to inclusive creativity and measurable results. It is a B Corp that values diversity, equity, inclusion, and belonging, though company size is not specified.
Write, optimize, and maintain complex SQL across multiple database platforms and Databricks.
Design and build data ingestion pipelines using API integrations, CDC patterns, and modern data engineering practices.
Develop and maintain Databricks Lakehouse architectures following the Medallion model.
Coforge is a global IT services company specializing in digital transformation and technology solutions. With over 20,000 employees worldwide, the company fosters a collaborative and innovative culture focused on delivering value to clients.
Design and implement standardized, normalized data models integrating heterogeneous source systems.
Own org-wide architectural decisions for data warehousing, orchestration, and pipeline reliability.
Build and maintain production-grade data services using Python, BigQuery, and dbt.
CodeRoad provides end-to-end software development services, helping businesses scale with ideal infrastructure solutions. The company offers nearshore technology services, including staff augmentation and dedicated IT teams, to empower businesses in an evolving digital landscape.
Design, build, and evolve the core data platform infrastructure including distributed query engines, orchestration, and warehousing.
Own the lakehouse infrastructure as code, managing deployments through Terraform and Ansible on Kubernetes.
Build and maintain low-latency streaming and CDC ingestion pipelines, as well as batch ingestion paths landing in Iceberg.
Alpaca is a US-headquartered global leader in agent-first brokerage infrastructure for stocks, ETFs, options, crypto, and more. The company has a global team of 400+ members, is backed by $400M in funding, and fosters a culture of curiosity, empathy, and accountability.
Design, build, and scale data pipelines and infrastructure in ClickHouse, Postgres, Python, and dbt.
Run the pipelines behind 500M+ labeled addresses, moving terabytes of streaming and batch data every day.
Own data quality, reliability, and observability end-to-end.
Nansen is an onchain analytics platform that provides insights to traders and funds. The company labels and tracks 500M+ blockchain addresses, fostering a culture of speed, ownership, and curiosity.
Deliver high-quality data sets by curating, consolidating, and manipulating large-scale data sources.
Build high-quality data pipelines and ETL processes on platforms like Snowflake and BigQuery.
Partner with analytics, data science, and machine learning teams to provide data solutions.
Tripadvisor connects people to experiences worth sharing, aiming to be the world's most trusted source for travel. With over 500 million reviews and 390 million monthly visitors, the company is data-driven and offers a unique, global work environment that captures the speed and innovation of a startup.
Design and build scalable data pipelines (batch & streaming) on GCP.
Develop and manage API-driven integrations (REST, JSON, event-based).
Work with healthcare data standards such as FHIR, HL7, and EDI.
Egen is a data-first consulting firm specializing in Google Cloud and Salesforce, helping clients drive action through data and insights. The company is fast-growing and entrepreneurial, with a culture focused on learning, problem-solving, and innovation.
Build and maintain data pipelines connecting source systems like e-commerce, marketing, and finance platforms to a data warehouse in Google BigQuery.
Model the data warehouse with clear logic, develop transformation models with tests and documentation, and ensure data reliability and consistent definitions.
Manage platform security, including access controls, pipeline monitoring, and automating processes with versioning, code reviews, and automated tests.
Naturtreu develops dietary supplements to help people care for their bodies consciously. The company is a 100% remote digital brand that has grown continuously since 2018, reaching hundreds of thousands of people monthly, with a team that values responsibility, practicality, and execution.
Collaborate with Data Science, Product Managers, and Software Engineers to build robust ETL pipelines for user-facing features.
Contribute to architecture decisions, observability tooling, and data quality initiatives to keep the platform robust.
Enforce engineering best practices across the AI/ML org, including code quality, testing, and documentation.
Federato is an AI-native platform for insurance, enabling insurers to provide affordable coverage for climate, cyber, and social inflation risks. It is a small, well-funded company backed by the investors behind Salesforce, Veeva, and Zoom, with a culture focused on first principles, learning, and fun.
Translate business goals into concrete data solutions.
Collaborate with peers, manage multiple priorities, and communicate project status.
Apply best practices in data pipelines, schemas, and system design.
DEPT is a growth invention company helping ambitious brands grow at the intersection of technology and marketing. With over 4,000 specialists, it values collaboration, curiosity, and getting things done.
Design and implement ETL pipelines using Apache Airflow, BigQuery, Python, and Spark to transform upstream data into curated data assets.
Provide technical leadership and best practices, mentoring other engineers and driving architecture decisions for high-performance systems.
Collaborate cross-functionally with Product Managers and end users to define key business questions and build relevant data sets.
InMarket is a leader in 360-degree consumer intelligence and real-time activation for top brands, offering a data-driven marketing platform. The company has a strong focus on technology and culture, with a commitment to diversity, equity, and inclusion, and offers competitive compensation and benefits.
Build new analyses and support existing ones using SQL and Python.
Apply software engineering principles like version control and continuous integration to the analytics codebase.
Expand our data warehouse with clean data ready for analysis.
Newsela is a leading education technology company dedicated to meaningful classroom learning for every student. It delivers integrated, AI-powered solutions to unlock student engagement and empower teachers.
Build and enhance data solutions and AI initiatives for insurance using GCP and BigQuery.
Code BigQuery procedures and implement scalable data models within the data lake.
Collaborate with cross-functional teams to deliver solutions aligned with business objectives and data governance guidelines.
Applied builds cloud software and AI-powered solutions for insurance agencies and brokers worldwide. With over 40 years of industry experience, the company fosters a people-first culture built on trust, inclusion, and growth.
Design and build batch data pipelines that ingest, validate, and transform multi-billion-row datasets.
Model complex real-world data including dimensional models and temporal data.
Develop and operate workloads on lakehouse platforms like Databricks, Spark, and Delta.
Simulmedia builds an advanced TV and streaming advertising platform. They have a team of engineers, data scientists, and designers who are obsessed with building cutting-edge technology.
Lead a high-impact data engineering team building the industry's first holistic real estate data engine.
Define and evolve the data platform strategy including ingestion, modeling, governance, and self-serve analytics.
Partner cross-functionally with Product, Analytics, and Engineering to ensure data enables customer and business outcomes.
Built is an AI-powered platform that connects and simplifies doing business in real estate, transforming how the industry finances, develops, and manages properties. They partner with over 350 lenders and 80,000 borrowers, powering 86,000 active projects valued at more than $300 billion.