Design production data pipelines using Databricks Lakeflow, Autoloader, and Structured Streaming with CI/CD.
Architect Lakehouse solutions including medallion architecture and Unity Catalog for analytics and AI.
Collaborate with product and backend engineers to design data models and APIs that serve the product.
Livefront helps companies design and build world-class digital products. They have worked with major brands and startups, fostering a culture built on respect, mutual trust, and egoless collaboration.
Design and drive the execution of a scalable, cloud-native target-state architecture for the data platform.
Define architectural patterns for scalable automated data ingestion and deliver hands-on proof-of-concepts using SQL, PySpark, and Serverless Python.
Drive governance, security, and engineering excellence across all platform components, ensuring alignment with business and product strategies.
G-P is a SaaS-based Global Employment Platform that enables clients to expand into over 180 countries. Our diverse, remote-first teams foster innovation and value every contribution, empowering employees with flexibility and resources.
Lead and deliver end-to-end Databricks implementation projects at enterprise scale.
Design and architect Lakehouse Platform solutions aligned with client business needs.
Apply deep Apache Spark expertise to optimize distributed workloads and runtime performance.
We are a data infrastructure and analytics consulting practice specializing in Databricks implementations for enterprise clients. Our team is composed of experienced engineers and architects who prioritize production-grade quality and continuous learning.
Lead the technical design and implementation of the RevOS data architecture within Databricks and the DHA War Data Platform.
Assess and restructure existing revenue-cycle data into a scalable, governed, and reusable data model aligned to the healthcare workflow.
Ensure the resulting data architecture directly supports production dashboards, KPI reporting, and drill-down analysis.
LMI is a digital solutions provider dedicated to accelerating government impact with innovation and speed. Headquartered in Tysons, Virginia, LMI serves the defense, space, healthcare, and energy sectors with a focus on agility and collaboration.
Design and deliver complex data platform components or migration solutions on AWS using Databricks.
Troubleshoot performance, scalability, and reliability issues in cloud-native data environments.
Communicate technical topics clearly to stakeholders and produce high-quality documentation.
Caylent is an AI-first cloud services company that helps organizations turn ambitious ideas into meaningful business impact. As a fully remote global company with employees in Canada, the United States, and Latin America, they celebrate diverse cultures and foster a community of technological curiosity.
You will work under the guidance of the Central Platform Team to build our Databricks engineering framework.
You will establish reusable development patterns and advance our core codebase.
You will use modern data engineering practices and AI-assisted tools to build scalable data solutions.
Talan is an international consulting group specializing in innovation and business transformation through technology. With over 7,200 consultants in 21 countries and a turnover of €850M, they are committed to delivering impactful, future-ready solutions.
Deliver complex data engineering projects for EMEA and US customers, individually or leading small teams.
Support pre-sales, mentor junior engineers, and contribute to hiring best-in-class data talent.
Work remotely in the UK with occasional travel to Budapest for onboarding and quarterly visits.
Datapao is a data engineering and cloud migration consulting company that helps clients solve complex data puzzles using technologies like Apache Spark and Databricks. They are a rapidly growing tech company with a high-transparency culture, emphasizing open feedback and no politics.
Design and build the next generation big data compute platform for ETL, analytics, and machine learning at Airbnb.
Operate, manage, and improve the reliability, performance, observability, and cost efficiency of the data platform.
Write maintainable, self-documenting code, perform code reviews, and contribute to open source software.
Airbnb is a global home-sharing platform that connects travelers with unique accommodations and experiences. The company has grown to over 5 million hosts and 2 billion guest arrivals, fostering a culture of belonging and innovation.
Design and ship data infrastructure at the core of modern data platforms, scalable pipelines, and reporting layers.
Work directly with the CEO and a small senior team, owning results from day one.
Build client-facing systems, translating business problems into architecture and delivering results.
Paradox Machines is a data and AI company that helps businesses make their data actually usable. They are a venture-backed company with a small, senior team that values ownership and impact.
Design and deliver data engineering solutions, including data lakehouse buildouts and cloud migrations, for EMEA and US customers.
Lead small delivery teams, mentor junior engineers, and support pre-sales processes to compound impact across the organization.
Work with Apache Spark/Databricks on AWS/Azure, manage customer relationships, and communicate technical concepts to non-technical audiences.
DATAPAO is a data engineering and cloud migration consultancy that helps customers solve complex data puzzles. The company emphasizes a high-transparency culture, open feedback, and a strong community, with a focus on innovation and impact.
Design, develop, and maintain scalable data pipelines and infrastructure.
Collaborate with cross-functional teams to ensure data integrity and usability.
Implement data quality and governance practices within data pipelines.
We help large enterprises transform with AI and tech-integrated solutions. With 8,000 employees across 25+ countries, we collaborate to build solutions with real impact.
Lead the design and evolution of scalable data, analytics, and AI platforms for global customer operations.
Architect and optimize Databricks-based data pipelines and enterprise data platforms for unified intelligence.
Drive development of ML, NLP, Generative AI, and operational intelligence capabilities to improve customer experience.
Ticketmaster is the world's largest ticket marketplace, connecting fans to live events globally. As part of Live Nation Entertainment, they employ thousands worldwide and foster an inclusive culture driven by teamwork and integrity.
Support new and existing customers in their data engineering needs, guiding them to make optimal technical decisions.
Build and operationalize complex data solutions, including data governance, security, and quality.
Collaborate across teams to deliver data products and educate end users on analytic environments.
SunnyData is a high-growth consulting company specializing in data engineering and AI, dedicated to the Databricks platform. They foster a collaborative and innovative culture with a focus on customer impact and career growth.
Develop and maintain data pipelines using Databricks, PySpark, Python, and SQL to transform raw financial data into reliable, curated datasets.
Integrate new data sources, support financial reconciliation, P&L routines, and global closing processes, ensuring data consistency and accuracy.
Collaborate with multidisciplinary teams to deliver high-quality data solutions, support AI initiatives, and implement continuous improvements in data performance and quality.
CI&T helps large companies transform AI potential into real business impact with AI deployment, AI-native execution, and tech-integrated business solutions. With 30 years of experience and 8,000 employees across 25+ countries, they collaborate to build solutions with real impact.
Design and optimize scalable data platforms, pipelines, and governance frameworks.
Lead data strategy and collaborate with leadership, engineering, and AI teams.
Mentor engineers and drive platform modernization with an AI-forward mindset.
Robots & Pencils is an applied AI engineering firm building AI co-workers for enterprise operations. Founded in 2009 with delivery centers in Canada, US, Eastern Europe, and Latin America, the company is a nimble alternative to traditional system integrators with teams averaging 15+ years of experience.
Develop, adjust, and evolve data pipelines using Databricks and PySpark.
Analyze processes migrated from legacy systems (SAS, Hive, Teradata) and make necessary adjustments.
Identify bottlenecks and implement performance improvements and optimizations in pipelines and large-scale data processing.
CI&T helps large companies transform AI potential into real business impact with AI deployment, AI-native execution, and tech-integrated business solutions. With 30 years of experience and 8,000 employees across 25+ countries, we accelerate innovation through expertise in agentic SDLC, application modernization, Data & AI, martech, and business strategy.
Lead the strategy and architecture of Alto's data platform, ensuring high-quality data fuels analytics and machine learning.
Manage and grow a high-performing data engineering team, fostering technical excellence and cross-functional partnership.
Drive data governance, reliability, and innovation in a regulated healthcare environment.
Fuze Health is a digital healthcare company that connects patients with care providers and resources through technology. It combines industry leaders including LetsGetChecked, Truepill, and Alto Pharmacy to create a unified force in healthcare.
Design and build reliable data pipelines using Spark, Kafka, Iceberg, and Airflow across batch, streaming, and real-time workloads.
Contribute to the evolution of the data lake and platform, including ingestion, processing, storage, and serving patterns.
Define and improve data quality, observability, reliability, and governance standards across data systems.
Webflow is an agentic web marketing platform for modern marketing teams, helping organizations build, manage, and optimize high-performing web experiences. The company is a growing, privately held organization that values grit, speed, and craft.
Translate legacy Databricks notebook logic into modern ELT patterns using Python and PySpark, ensuring data contract preservation and reverse-view strategies.
Build scalable ingestion pipelines with YAML configurations and orchestrate automated DAG generation via workflow schedulers, applying quality assertions and migration validation.
Collaborate with business Data Stewards to align dependencies, negotiate refactoring scope, and validate migrated outputs for production systems.
CI&T helps large companies transform AI potential into real business impact with AI deployment, AI-native execution, and tech-integrated business solutions. With 30 years of tech transformation experience and over 8,000 CI&Ters in 25+ countries, we accelerate innovation through agentic SDLC, application modernization, Data & AI, martech, and business strategy.
Build and optimize scalable data pipelines using Databricks, Spark, and PySpark for cloud migration.
Design and maintain data models (Bronze, Silver, Gold layers) to enable analytical insights.
Ensure data quality and governance while managing the technical roadmap and communicating with stakeholders.
CI&T helps large companies transform AI potential into real business impact with AI deployment and tech-integrated solutions. They have 8,000 employees across 25+ countries, fostering a collaborative and innovative culture.