Develop and maintain scalable, resilient data pipelines using PySpark and distributed processing frameworks.
Modernize legacy processes and design data solutions on AWS, building data products and complex transformations.
Support data quality and optimization initiatives, collaborating with stakeholders to translate requirements into technical solutions.
Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. It focuses on efficiency and fairness in recruitment, promoting diversity and inclusion in the workplace.
You will develop and maintain end-to-end data pipelines and contribute to Samsara's Data Platform for advanced automation and analytics.
You will design, build, and optimize large-scale Spark and PySpark workflows for batch and streaming data processing.
You will build and maintain MCP servers and AI agents, and champion data engineering best practices across the team.
Samsara is the pioneer of the Connected Operations Cloud, enabling organizations to harness IoT data for actionable insights. As a recently public company, they foster a culture of autonomy, support, and rapid career development in a hyper-growth environment.
Own the identity data model and coordinate cross-ecosystem mappings for billions of package observations.
Architect streaming and batch pipelines using Java, Spark, HBase, Databricks, and AWS primitives.
Lead design of data contracts for downstream teams, ensuring schema stability and versioning.
Sonatype is the software supply chain security company, providing end-to-end security solutions including proactive protection against malicious open source and enterprise-grade SBOM management. With over 2,000 organizations including 70% of the Fortune 100 and 15 million developers relying on its platform, Sonatype is a market leader in software supply chain security.
Design, build, and maintain scalable data pipelines and ETL/ELT processes using Databricks and Spark.
Architect and optimize data models and storage solutions for analytics and operational use.
Implement observability, alerting, and data quality monitoring for critical pipelines.
Sonatype provides end-to-end software supply chain security solutions, protecting against malicious open source and managing SBOMs. With over 2,000 organizations and 15 million developers using its platform, it focuses on innovation and security in software development.
Design and build scalable data pipelines using Apache Spark and cloud technologies.
Develop and maintain data integrations and transformation processes from multiple sources.
Collaborate with technical teams to ensure data quality and optimize data architecture.
Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. They focus on using technology to streamline the hiring process and offer a remote work environment.
Design and build batch data pipelines that ingest, validate, and transform multi-billion-row datasets.
Model complex real-world data including dimensional models and temporal data.
Develop and operate workloads on lakehouse platforms like Databricks, Spark, and Delta.
Simulmedia builds an advanced TV and streaming advertising platform. They have a team of engineers, data scientists, and designers who are obsessed with building cutting-edge technology.
Design and implement a petabyte-scale AWS data platform using advanced engineering and cloud-native practices.
Build and optimize highly scalable data pipelines using JVM languages, Spark, and Python for production-ready data processing.
Act as a principal-level technical leader, mentoring engineers and driving architectural decisions to guide platform evolution.
Experian is a global data and technology company that powers opportunities for people and businesses around the world. With 25,200 employees across 32 countries, they have a people-centric, inclusive culture recognized as a World's Best Workplace.
Design end-to-end cloud-native data architectures using AWS and Databricks.
Architect scalable batch and real-time data processing solutions using Apache Kafka and streaming technologies.
Collaborate with business stakeholders to translate requirements into scalable technical solutions.
Allwyn Corp is a technology company that provides data solutions and cloud-native services. The company values innovation and technical excellence, fostering a culture of collaboration and continuous learning.
Define and enforce best practices and coding standards across the project.
Design, develop, and maintain robust and scalable Spark applications.
Work closely with cross-functional teams to deliver high-quality software solutions.
Exadel is an AI-first global tech company with 25+ years of engineering leadership and over 2,000 team members. With a culture of ambition, collaboration, and constant evolution, they power Fortune 500 clients including HBO, Microsoft, Google, and Starbucks.
Design and implement scalable data architectures to meet business needs.
Develop and optimize data pipelines for ingesting and processing large volumes of data.
Provide technical leadership and mentorship to junior engineers.
Oportun is a mission-driven financial services company that provides affordable credit and financial tools. Since inception, it has provided over $21.3 billion in credit and saved members $2.5 billion, and fosters a diverse, inclusive culture.
Maintain and monitor Kafka, Hadoop, Presto, and RDBMS systems.
Ingest, validate, and process internal and third-party data flows.
Build Kafka consumers using Spark Streaming for near-real-time aggregation.
PulsePoint sits at the intersection of healthcare and adtech, helping brands interpret health journey signals. We are 300+ employees, growing, and a leading player in the US healthcare ad market.
Design, develop, and maintain robust backend services, APIs, and data integration layers for a fintech platform.
Write clean, efficient, and maintainable code following strict coding standards and participate in code reviews.
Optimize database schemas, improve testing infrastructure, and enhance monitoring and alerting capabilities.
Ryz Labs is a talent platform that connects skilled engineers with leading companies. They focus on building high-performing remote teams and fostering a culture of collaboration and growth.
Design, develop, maintain, and optimize ETL and data transformation processes.
Develop and maintain Apache Spark-based data pipelines and contribute to data integration initiatives.
Ensure data quality, reliability, and performance across data processing solutions.
Talan is an international consulting group specializing in innovation and business transformation through technology. With over 7,200 consultants in 21 countries and a turnover of €850M, the company delivers impactful, future-ready solutions.
Design and deliver complex data platform components or migration solutions on AWS using Databricks.
Troubleshoot performance, scalability, and reliability issues in cloud-native data environments.
Communicate technical topics clearly to stakeholders and produce high-quality documentation.
Caylent is an AI-first cloud services company that helps organizations turn ambitious ideas into meaningful business impact. As a fully remote global company with employees in Canada, the United States, and Latin America, they celebrate diverse cultures and foster a community of technological curiosity.
Lead and deliver end-to-end Databricks implementation projects at enterprise scale.
Design and architect Lakehouse Platform solutions aligned with client business needs.
Apply deep Apache Spark expertise to optimize distributed workloads and runtime performance.
We are a data infrastructure and analytics consulting practice specializing in Databricks implementations for enterprise clients. Our team is composed of experienced engineers and architects who prioritize production-grade quality and continuous learning.
Design and implement ETL pipelines using Apache Airflow, BigQuery, Python, and Spark to transform upstream data into curated data assets.
Provide technical leadership and best practices, mentoring other engineers and driving architecture decisions for high-performance systems.
Collaborate cross-functionally with Product Managers and end users to define key business questions and build relevant data sets.
InMarket is a leader in 360-degree consumer intelligence and real-time activation for top brands, offering a data-driven marketing platform. The company has a strong focus on technology and culture, with a commitment to diversity, equity, and inclusion, and offers competitive compensation and benefits.
Design and implement the UniForm write layer (Delta + Iceberg dual metadata).
Build GCS → BigQuery ingestion pipelines for structured operational datasets.
Develop and implement in Python and Spark, including Kafka-based CDC patterns for real-time ingestion.
Railroad19, Inc develops customized software solutions and provides software development services. They are a specialized team of developers and architects with a culture of hard work and industry leadership, valuing employees and offering remote work from the US.
Architect and develop robust, end-to-end machine learning solutions, managing the complete ML lifecycle from development to production.
Collaborate closely with data engineers and data scientists to create highly scalable solutions and integrate them across the software development life cycle.
Evaluate cloud solutions for performance and cost-efficiency, and communicate technical concepts clearly to stakeholders and teams.
Coderio designs and delivers scalable digital solutions for global companies, combining strong technical expertise with a product mindset to lead complex software initiatives end-to-end. They work with international clients, value autonomy and clear communication, and build long-term partnerships through technical excellence.
Design, develop, and maintain scalable, production-ready data pipelines and data products using Spark (Python/SQL) in a Databricks environment.
Lead the integration and transformation of complex data from diverse DoD and federal health systems into reliable, reusable data products.
Provide technical guidance and mentorship to other engineers, helping teams navigate complex technical challenges.
540 is a forward-thinking company that delivers innovative technology solutions for government missions. The team has a culture of breaking down barriers and solving mission-critical problems.
Guide strategic enterprise customers through cloud data engineering transformations, including performance testing and production-ready pipeline architecture.
Prove platform value by architecting solutions for big data, data warehousing, and lakehouse use cases.
Support technical sales through custom proofs of concept, workload sizing, and community workshops.
Databricks is the Data and AI company, providing a unified platform for data, analytics, and AI to over 20,000 organizations worldwide, including 70% of the Fortune 500. Headquartered in San Francisco with 30+ offices globally, it fosters a diverse and inclusive culture and offers comprehensive benefits.