Own the Airflow codebase end-to-end, building reusable templates and enforcing standards.
Build and extend large-scale Spark pipelines on AWS Glue and support migration to Databricks.
Drive data model improvements around commercial pharma data with focus on structure and lineage.
Veeva Systems is a mission-driven organization and pioneer in industry cloud, helping life sciences companies bring therapies to patients faster. As one of the fastest-growing SaaS companies in history, it surpassed $3B in revenue and values doing the right thing, customer success, employee success, and speed.
You will develop and maintain end-to-end data pipelines and contribute to Samsara's Data Platform for advanced automation and analytics.
You will design, build, and optimize large-scale Spark and PySpark workflows for batch and streaming data processing.
You will build and maintain MCP servers and AI agents, and champion data engineering best practices across the team.
Samsara is the pioneer of the Connected Operations Cloud, enabling organizations to harness IoT data for actionable insights. As a recently public company, they foster a culture of autonomy, support, and rapid career development in a hyper-growth environment.
Architect enterprise-scale Databricks data platforms with guardrails, automation, and standards to enable safe self-service by business teams.
Design high-volume batch and real-time streaming pipelines using PySpark and Structured Streaming, integrating Unity Catalog with third-party metadata.
Communicate architectural decisions to technical and business stakeholders, produce documentation, and partner with client teams to translate use-case needs into platform capabilities.
Caylent is an AI-first cloud services company that helps organizations turn ambitious ideas into meaningful business impact using AWS, AI, and Anthropic's Claude platform. They are a fully remote global company with employees across Canada, the US, and Latin America, fostering a culture of technological curiosity and inclusion.
Guide strategic enterprise customers through cloud data engineering transformations, including performance testing and production-ready pipeline architecture.
Prove platform value by architecting solutions for big data, data warehousing, and lakehouse use cases.
Support technical sales through custom proofs of concept, workload sizing, and community workshops.
Databricks is the Data and AI company, providing a unified platform for data, analytics, and AI to over 20,000 organizations worldwide, including 70% of the Fortune 500. Headquartered in San Francisco with 30+ offices globally, it fosters a diverse and inclusive culture and offers comprehensive benefits.
Build, expand, and optimize data infrastructure to create the most accurate dataset of identities and their relationships.
Develop and operate secure, scalable, and reliable data ingestion and ETL/ELT pipelines that meet product requirements.
Design and maintain a data observability framework to ensure data meets strict quality and freshness standards.
SentiLink provides innovative identity and risk solutions, empowering institutions and individuals to transact with confidence. The company is growing rapidly, has verified hundreds of millions of identities, and is backed by top investors like Craft Ventures and Andreessen Horowitz, with offices across the US and India.
Architect and evolve our data platform using Snowflake and dbt.
Design and implement batch and streaming pipelines in AWS.
Establish best practices for data modeling, governance, and security.
BlueMatrix develops web applications for the authoring, distribution, and analysis of investment research. The team consists of ambitious professionals with offices across North America and Europe.
Design, build, and maintain serverless data pipelines on AWS using Lambda, Glue, Athena, S3, and Step Functions to support analytics and AI.
Develop ETL/ELT workflows to ingest, cleanse, and load data from internal and third-party sources into a data lake, ensuring data quality and reliability.
Collaborate with Analytics, Data Science, and AI teams to understand data requirements and create robust data models and feature stores.
ParetoHealth is redefining how employers fund healthcare as the largest and fastest-growing benefits captive in the United States. Headquartered in Philadelphia, it is a rapidly growing company driven by core values such as Fire in the Belly and For the Greater Good.
Design and maintain data pipelines on AWS using Spark, Python, and SQL.
Orchestrate workflows with Airflow, containerize with Docker and ECS.
Collaborate within agile teams to deliver high-quality data solutions.
Devoteam is a leading European consulting firm focused on digital strategy, technology platforms, cybersecurity, and business transformation through technology. With over 10,000 employees across 20 countries in Europe, the Middle East, and Africa, they foster a close-knit professional culture.
Design and implement scalable data architectures to meet business needs.
Develop and optimize data pipelines for ingesting and processing large volumes of data.
Provide technical leadership and mentorship to junior engineers.
Oportun is a mission-driven financial services company that provides affordable credit and financial tools. Since inception, it has provided over $21.3 billion in credit and saved members $2.5 billion, and fosters a diverse, inclusive culture.
Design, build, and maintain scalable data pipelines and platform capabilities to power analytics, AI/ML, and healthcare products.
Implement data quality, governance, and security controls, ensuring healthcare data is handled securely and in compliance.
Provide technical leadership, mentor engineers, and collaborate across teams to translate requirements into scalable data solutions.
Experity is a mission-driven team transforming on-demand healthcare across the U.S., empowering urgent care clinics with industry-leading software. They foster a culture of care, growth, and celebration with day-one benefits, career development, and a supportive team environment.
Design, build, and maintain scalable data pipelines using Python and AWS DynamoDB.
Develop and optimize DynamoDB data models and indexing strategies for high throughput.
Establish data validation, monitoring, and alerting mechanisms to ensure pipeline reliability.
Kunai builds full-stack technology solutions for banks, credit and payment networks, infrastructure providers, and their customers. With a team that thrives in a culture of collaboration and creativity, the company offers competitive compensation and professional development opportunities.
Become a trusted data and AI advisor, translating business questions into AI-ready data architectures.
Design and implement AI-optimized data platforms, including cloud data warehouses, lakehouses, and ETL/ELT pipelines.
Engineer modern ELT/ETL pipelines and data models using SQL, Python, and tools like Snowflake, Databricks, and dbt.
Aimpoint Digital is a fully remote data and analytics consultancy that partners with innovative software providers to solve complex business problems. The team is dynamic and collaborative, working independently on client engagements across industries.
Lead the design and implementation of complex data pipelines and infrastructure across multiple domains
Mentor junior engineers and raise team-wide engineering standards
Optimize and monitor production workflows using tools like Airflow, dbt, and AWS
Cohere Health's clinical intelligence platform connects health plans and providers to optimize care. We are a growing company backed by leading investors, with a culture of empathy and collaboration.
Develop and maintain scalable, resilient data pipelines using PySpark and distributed processing frameworks.
Modernize legacy processes and design data solutions on AWS, building data products and complex transformations.
Support data quality and optimization initiatives, collaborating with stakeholders to translate requirements into technical solutions.
Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. It focuses on efficiency and fairness in recruitment, promoting diversity and inclusion in the workplace.
Develop and maintain our AWS-based data platform for a quickly scaling business.
Build and own our data warehouse and the pipelines that feed it from product, CRM and third-party systems.
Ensure data quality, performance and scalability and own data protection obligations within our pipelines.
OneDome is a UK-based property technology company on a mission to make buying, selling and financing a home simpler and faster. We are a small, quickly scaling company with a distributed team across Ukraine and Azerbaijan.
Design and implement scalable data ingestion and integration pipelines for structured and unstructured data, ensuring AI-ready and governed data.
Build and scale retrieval infrastructure including vector storage, embedding pipelines, and hybrid search.
Lead pragmatic platform evolution by defining clear contracts between AI services and data systems, improving developer experience.
Caseware is a Canadian Fintech company that has led the global audit and accounting software industry for over 30 years, with over 500,000 users across 130 countries. They foster a culture of inclusion and innovation, with a commitment to 'Many Voices, One Team'.
Design, build, and maintain scalable data pipelines on AWS using S3, Glue, Redshift, Athena, and EMR.
Provision and manage AWS data infrastructure using Terraform and CloudFormation, and orchestrate workflows with Apache Airflow.
Implement HIPAA and HITRUST security controls, and collaborate with teams to ensure data quality and compliance.
AspenView Technology Partners specializes in building high-performing nearshore IT teams for North American clients. They are a people-first company with a culture that blends U.S. innovation and Latin American talent, offering flexible work and growth opportunities.
Design and build batch data pipelines that ingest, validate, and transform multi-billion-row datasets.
Model complex real-world data including dimensional models and temporal data.
Develop and operate workloads on lakehouse platforms like Databricks, Spark, and Delta.
Simulmedia builds an advanced TV and streaming advertising platform. They have a team of engineers, data scientists, and designers who are obsessed with building cutting-edge technology.
Build and own data pipelines, models, and infrastructure that turn raw labor market data into systems connecting people to jobs.
Shape data modeling in the warehouse, ensure trustworthy analytics and reporting, and build pipelines for matching and recommendation models.
Partner closely with Engineering, Product, and VP of Data & AI to decide how the platform is built.
FutureFit AI is an AI-powered platform that helps people get better jobs faster, focusing on those facing barriers to opportunity. The company has 30-50 employees across the US and Canada, with a culture of high trust, high impact, and a will to win.
Build and maintain batch data pipelines ingesting from Workday and legacy systems into a medallion lakehouse on Azure Databricks and Microsoft Fabric.
Develop transformation logic across Bronze/Silver/Gold layers using Spark, Python, and SQL, and implement data-quality checks.
Work within a federated architecture and Unity Catalog security model, contributing to documentation and code reviews.
CampusWorks is a consulting firm that helps higher education institutions leverage technology and managed services to drive transformation and student success. Founded in 1999, it is a large, virtual company with a culture that values work-life balance and employee appreciation.