Design, build, and maintain scalable data pipelines and systems using Python, PySpark, and AWS-native services.
Develop and optimize data platforms on AWS, leveraging technologies like Redshift, Glue, and Athena.
Mentor junior engineers and collaborate with cross-functional teams in an Agile environment.
Capital Technology Group provides expert consulting services in software development, digital transformation, and data analytics. Recognized as a Top Workplace by The Washington Post, they foster a culture where employees feel trusted and empowered.
Design and build scalable data pipelines using Apache Spark and cloud technologies.
Develop and maintain data integrations and transformation processes from multiple sources.
Collaborate with technical teams to ensure data quality and optimize data architecture.
Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. They focus on using technology to streamline the hiring process and offer a remote work environment.
Own the Airflow codebase end-to-end, building reusable templates and enforcing standards.
Build and extend large-scale Spark pipelines on AWS Glue and support migration to Databricks.
Drive data model improvements around commercial pharma data with focus on structure and lineage.
Veeva Systems is a mission-driven organization and pioneer in industry cloud, helping life sciences companies bring therapies to patients faster. As one of the fastest-growing SaaS companies in history, it surpassed $3B in revenue and values doing the right thing, customer success, employee success, and speed.
Design and implement scalable data architectures to meet business needs.
Develop and optimize data pipelines for ingesting and processing large volumes of data.
Provide technical leadership and mentorship to junior engineers.
Oportun is a mission-driven financial services company that provides affordable credit and financial tools. Since inception, it has provided over $21.3 billion in credit and saved members $2.5 billion, and fosters a diverse, inclusive culture.
Structure, integrate, and maintain scalable data solutions for credit origination, monitoring, and risk analysis.
Build and evolve the credit analytics layer including datamarts, indicators, and reusable components.
Implement data quality, governance, and observability controls while ensuring compliance with LGPD.
Jobgether is an AI-powered recruitment platform that connects candidates with job opportunities using an objective matching process. It operates with a small team, focusing on efficiency, fairness, and data privacy.
Own dashboard changes end to end in QuickSight and ThoughtSpot, including new metrics, SPICE refresh, and post-release validation.
Build and maintain batch and streaming ETL pipelines on AWS using Glue (PySpark), S3, Redshift, and Airflow DAGs.
Investigate customer-reported analytics discrepancies, root-cause data issues, and communicate findings across teams.
Eltropy is a digital conversations platform for credit unions and community financial institutions in the US. The Data Engineering & Analytics team builds AWS data pipelines and customer-facing dashboards that power analytics; they seek a collaborative engineer who learns fast and works across teams.
Design, develop, and maintain scalable ETL/ELT pipelines and data integration processes.
Build and optimize cloud-based data architectures to support analytics and business intelligence initiatives.
Ensure data quality, consistency, governance, and reliability through validation, monitoring, and automated quality checks.
CI&T helps large enterprises transform AI potential into business impact with AI deployment, AI-native execution, and tech-integrated solutions. With 30 years of experience and 8,000 employees across 25 countries, they collaborate to build solutions with real impact.
Lead impactful customer technical projects by delivering production-grade systems spanning data engineering, AI, and application development.
Guide strategic customers in implementing end-to-end big data and AI projects, including architecture, design, build, and deployment.
Empower customers by providing architecture guidance and ensuring solutions are secure, scalable, and aligned with Databricks best practices.
Databricks is the data and AI company. More than 10,000 organizations worldwide, including Comcast and Condé Nast, rely on the Databricks Data Intelligence Platform to unify and democratize data, analytics and AI. The company is headquartered in San Francisco and fosters a diverse, inclusive culture.
You will develop and maintain end-to-end data pipelines and contribute to Samsara's Data Platform for advanced automation and analytics.
You will design, build, and optimize large-scale Spark and PySpark workflows for batch and streaming data processing.
You will build and maintain MCP servers and AI agents, and champion data engineering best practices across the team.
Samsara is the pioneer of the Connected Operations Cloud, enabling organizations to harness IoT data for actionable insights. As a recently public company, they foster a culture of autonomy, support, and rapid career development in a hyper-growth environment.
Lead the modernization of the data platform by migrating legacy Hadoop, Spark, and Impala pipelines to a scalable Databricks architecture.
Accelerate migration efforts using AI coding assistants like Claude and Codex to convert SQL and modernize ETL workflows.
Design and optimize scalable data pipelines with Airflow and Databricks, ensuring reliability, cost-efficiency, and data quality.
LivePerson is a leader in enterprise conversational AI and digital transformation, powering nearly a billion conversational interactions monthly. Fast Company named them the #1 Most Innovative AI Company, and they foster a remote-first, innovative culture.
Own production data pipelines end-to-end, ensuring reliability, performance, and scalability. - Build and optimize data solutions using PySpark and Foundry tools to process massive datasets. - Champion data quality by designing monitoring and validation frameworks to guarantee trusted data.
We are an IT solutions company providing innovative information and communication technology services. We have grown to over 3,900 employees and are the second largest employer in eastern Slovakia, with a culture focused on continuous improvement and transformation.
Build, expand, and optimize data infrastructure to create the most accurate dataset of identities and their relationships.
Develop and operate secure, scalable, and reliable data ingestion and ETL/ELT pipelines that meet product requirements.
Design and maintain a data observability framework to ensure data meets strict quality and freshness standards.
SentiLink provides innovative identity and risk solutions, empowering institutions and individuals to transact with confidence. The company is growing rapidly, has verified hundreds of millions of identities, and is backed by top investors like Craft Ventures and Andreessen Horowitz, with offices across the US and India.
Design and build batch-oriented data pipelines and ETL workflows using AWS Glue, Lambda, and Step Functions.
Develop and optimize ingestion pipelines and analytical data models with Athena, Redshift, and ClickHouse.
Collaborate with backend engineers to integrate data workflows with microservices and ensure data quality and observability.
SavvyMoney is a US-based financial technology company providing integrated credit score and personal finance solutions to over 1,600 bank and credit union partners. It was recognized as a top workplace and Inc. 5000 fastest growing company, with a distributed team across the US, Canada, Europe, and India.
Build and maintain batch data pipelines ingesting from Workday and legacy systems into a medallion lakehouse on Azure Databricks and Microsoft Fabric.
Develop transformation logic across Bronze/Silver/Gold layers using Spark, Python, and SQL, and implement data-quality checks.
Work within a federated architecture and Unity Catalog security model, contributing to documentation and code reviews.
CampusWorks is a consulting firm that helps higher education institutions leverage technology and managed services to drive transformation and student success. Founded in 1999, it is a large, virtual company with a culture that values work-life balance and employee appreciation.
Design and deploy highly performant end-to-end data architectures using Databricks and Apache Spark.
Develop and manage scalable batch and streaming solutions in cloud ecosystems like AWS, Azure, or GCP.
Lead large-scale data engineering projects, ensuring technical delivery and client management.
Derex Technologies Inc provides IT consulting, staffing solutions, and software services to global clients. With over two decades of experience, the company delivers high-quality technology professionals across various industries.
Develop highly scalable and reliable data systems on AWS cloud platform for data-centric products.
Collaborate with Data Science teams to incorporate Machine Learning algorithms using Python and data engineering tools.
Lead and mentor team members, contribute to Agile cycles, and ensure timely delivery of documented software.
Experian is a global data and technology company that powers opportunities for people and businesses worldwide. With 22,500 employees across 32 countries, they foster a culture of innovation and data-driven solutions.
Own and evolve notebooks across the full medallion architecture, from raw ingestion through a fully modeled dimensional layer.
Design and implement dimensional modeling artifacts such as facts, dimensions, and slowly changing dimensions for downstream business intelligence.
Serve as the primary technical reference, reviewing pull requests, mentoring team members, and proposing architectural improvements.
CI&T helps large enterprises transform the potential of AI into real business impact with AI Deployment, AI-native execution, and tech-integrated business solutions. With 30 years of experience and 8,000 employees across 25 countries, we foster a collaborative culture that accelerates innovation in Agentic SDLC, Application modernization, Data & AI, Martech, and Business strategy.
Guide strategic enterprise customers through cloud data engineering transformations, including performance testing and production-ready pipeline architecture.
Prove platform value by architecting solutions for big data, data warehousing, and lakehouse use cases.
Support technical sales through custom proofs of concept, workload sizing, and community workshops.
Databricks is the Data and AI company, providing a unified platform for data, analytics, and AI to over 20,000 organizations worldwide, including 70% of the Fortune 500. Headquartered in San Francisco with 30+ offices globally, it fosters a diverse and inclusive culture and offers comprehensive benefits.
Design and build batch data pipelines that ingest, validate, and transform multi-billion-row datasets.
Model complex real-world data including dimensional models and temporal data.
Develop and operate workloads on lakehouse platforms like Databricks, Spark, and Delta.
Simulmedia builds an advanced TV and streaming advertising platform. They have a team of engineers, data scientists, and designers who are obsessed with building cutting-edge technology.
Design, develop, and maintain scalable ETL/ELT pipelines
Build and optimize data processing solutions using Python and Apache Spark
Develop and support real-time data streaming applications using Kafka
Sigma Software develops modern digital solutions for global businesses. They foster an international remote-first environment where engineers can grow, innovate, and influence technical decisions.