Design and build scalable data pipelines using Apache Spark and cloud technologies.
Develop and maintain data integrations and transformation processes from multiple sources.
Collaborate with technical teams to ensure data quality and optimize data architecture.
Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. They focus on using technology to streamline the hiring process and offer a remote work environment.
Design and deliver near real-time data solutions for the analytics platform.
Analyze business needs, optimize data models, and identify slow queries for performance improvement.
Write clean, scalable code using Scala, Python, and SQL while mentoring team members.
Aircall is an AI-powered customer communications platform used by 22,000+ companies worldwide. It is a unicorn startup with offices across multiple countries and a focus on innovation and collaboration.
Design and build batch data pipelines that ingest, validate, and transform multi-billion-row datasets.
Model complex real-world data including dimensional models and temporal data.
Develop and operate workloads on lakehouse platforms like Databricks, Spark, and Delta.
Simulmedia builds an advanced TV and streaming advertising platform. They have a team of engineers, data scientists, and designers who are obsessed with building cutting-edge technology.
Design end-to-end cloud-native data architectures using AWS and Databricks.
Architect scalable batch and real-time data processing solutions using Apache Kafka and streaming technologies.
Collaborate with business stakeholders to translate requirements into scalable technical solutions.
Allwyn Corp is a technology company that provides data solutions and cloud-native services. The company values innovation and technical excellence, fostering a culture of collaboration and continuous learning.
Design, develop, and maintain scalable ETL/ELT pipelines and data integration processes.
Build and optimize cloud-based data architectures to support analytics and business intelligence initiatives.
Ensure data quality, consistency, governance, and reliability through validation, monitoring, and automated quality checks.
CI&T helps large enterprises transform AI potential into business impact with AI deployment, AI-native execution, and tech-integrated solutions. With 30 years of experience and 8,000 employees across 25 countries, they collaborate to build solutions with real impact.
You will develop and maintain end-to-end data pipelines and contribute to Samsara's Data Platform for advanced automation and analytics.
You will design, build, and optimize large-scale Spark and PySpark workflows for batch and streaming data processing.
You will build and maintain MCP servers and AI agents, and champion data engineering best practices across the team.
Samsara is the pioneer of the Connected Operations Cloud, enabling organizations to harness IoT data for actionable insights. As a recently public company, they foster a culture of autonomy, support, and rapid career development in a hyper-growth environment.
Design, build, and maintain data infrastructure and pipelines spanning both batch and real-time workloads.
Build and maintain scalable ETL/ELT pipelines using Python, SQL, Spark, and orchestration frameworks.
Partner with data scientists, ML engineers, and product teams to deliver data products and establish data quality frameworks.
Gemini is a global crypto and Web3 platform founded by Cameron and Tyler Winklevoss in 2014, offering a wide range of crypto products to individuals and institutions in over 70 countries. As a publicly traded company, Gemini is poised to accelerate the vision of reshaping the global financial system with greater scale, reach, and impact.
Design and migrate scalable data solutions using modern cloud technologies like AWS and Data Mesh.
Convert Alteryx workflows to AWS-native architectures such as AWS Glue and improve performance.
Build and maintain reliable data pipelines using distributed processing tools like Spark and Airflow.
The partner company provides data engineering and cloud solutions aimed at helping organizations modernize their data platforms. They are a collaborative team of technology professionals focused on innovation and scalable architectures.
Build, expand, and optimize data infrastructure to create the most accurate dataset of identities and their relationships.
Develop and operate secure, scalable, and reliable data ingestion and ETL/ELT pipelines that meet product requirements.
Design and maintain a data observability framework to ensure data meets strict quality and freshness standards.
SentiLink provides innovative identity and risk solutions, empowering institutions and individuals to transact with confidence. The company is growing rapidly, has verified hundreds of millions of identities, and is backed by top investors like Craft Ventures and Andreessen Horowitz, with offices across the US and India.
Own dashboard changes end to end in QuickSight and ThoughtSpot, including new metrics, SPICE refresh, and post-release validation.
Build and maintain batch and streaming ETL pipelines on AWS using Glue (PySpark), S3, Redshift, and Airflow DAGs.
Investigate customer-reported analytics discrepancies, root-cause data issues, and communicate findings across teams.
Eltropy is a digital conversations platform for credit unions and community financial institutions in the US. The Data Engineering & Analytics team builds AWS data pipelines and customer-facing dashboards that power analytics; they seek a collaborative engineer who learns fast and works across teams.
Design and build scalable batch data pipelines using AWS Glue, Lambda, and Step Functions.
Optimize analytical data models and query layers with Amazon Athena, Redshift, and ClickHouse.
Collaborate with backend engineers to integrate data workflows with microservices and event-driven architectures.
The company is a technology organization focused on building scalable data platforms and analytics solutions. They operate with a distributed international team and emphasize modern engineering practices.
Support Snowflake-to-Databricks data migration activities.
Design and implement scalable data pipelines using Snowflake, Databricks, SQL, Python, and dbt.
Optimize data workflows for performance, scalability, cost efficiency, and maintainability.
Hiflylabs is a Budapest-based company delivering Data Warehouses, Business Intelligence, and Data Analytics solutions. With over 250 employees and 10 years of experience, they serve financial, telecommunication, and energy sectors, fostering a culture of innovation and collaboration.
Own the Airflow codebase end-to-end, building reusable templates and enforcing standards.
Build and extend large-scale Spark pipelines on AWS Glue and support migration to Databricks.
Drive data model improvements around commercial pharma data with focus on structure and lineage.
Veeva Systems is a mission-driven organization and pioneer in industry cloud, helping life sciences companies bring therapies to patients faster. As one of the fastest-growing SaaS companies in history, it surpassed $3B in revenue and values doing the right thing, customer success, employee success, and speed.
Design, build, and maintain scalable data pipelines and systems using Python, PySpark, and AWS-native services.
Develop and optimize data platforms on AWS, leveraging technologies like Redshift, Glue, and Athena.
Mentor junior engineers and collaborate with cross-functional teams in an Agile environment.
Capital Technology Group provides expert consulting services in software development, digital transformation, and data analytics. Recognized as a Top Workplace by The Washington Post, they foster a culture where employees feel trusted and empowered.
Design and maintain scalable batch and near real-time data pipelines in Databricks, integrating data from various sources.
Build unified customer profiles and support identity resolution, deduplication, enrichment, and Golden Record creation.
Enable CRM and marketing teams to access trusted, activation-ready customer data with minimal latency.
Massive Rocket is a rapidly scaling Braze and Snowflake agency that transforms how digital marketing, product, and engineering teams connect. They have grown quickly in five years and are aiming for $100M in revenue, with a culture of ownership, collaboration, and growth.
Extract, clean, transform, and store data required for use cases.
Design and manage data integration pipelines, developing high-quality ETLs and advanced analytics tools.
Monitor pipeline stability, run quality controls, and optimize costs, performance, and data security.
Redbee is a technology company with over 14 years of experience redefining how digital products are built in the financial industry. They co-create innovative solutions with clients, focusing on delivering fast, quality value through scalable and strategic products.
Collaborate with Data Science, Product Managers, and Software Engineers to build robust ETL pipelines for user-facing features.
Contribute to architecture decisions, observability tooling, and data quality initiatives to keep the platform robust.
Enforce engineering best practices across the AI/ML org, including code quality, testing, and documentation.
Federato is an AI-native platform for insurance, enabling insurers to provide affordable coverage for climate, cyber, and social inflation risks. It is a small, well-funded company backed by the investors behind Salesforce, Veeva, and Zoom, with a culture focused on first principles, learning, and fun.
Own the design, delivery, and operation of major data products including ETL pipelines and storage solutions.
Partner with engineering, product, and business stakeholders to define solutions and drive projects from design to production.
Mentor engineers, lead technical strategy, and participate in on-call rotation for production systems.
Lime is the largest global shared micromobility company, on a mission to make transportation shared, affordable, and carbon-free. It has powered over one billion rides across 30 countries and is a Time 100 Most Influential Company.
Build and maintain scalable data pipelines using Python, dbt, and Snowflake-native tooling on AWS.
Develop and maintain Snowflake semantic views to support analytics and reporting needs.
Implement data access controls and ensure pipelines meet enterprise data governance standards.
The Enterprise Analytics Center of Excellence is building a next-generation analytics stack centered on Snowflake, AWS, and Sigma. The team focuses on collaboration, continuous improvement, and enabling analytics teams with reliable data.