Translate legacy Databricks notebook logic into modern ELT patterns using Python and PySpark, ensuring data contract preservation and reverse-view strategies.
Build scalable ingestion pipelines with YAML configurations and orchestrate automated DAG generation via workflow schedulers, applying quality assertions and migration validation.
Collaborate with business Data Stewards to align dependencies, negotiate refactoring scope, and validate migrated outputs for production systems.
Develop, adjust, and evolve data pipelines using Databricks and PySpark.
Analyze processes migrated from legacy systems (SAS, Hive, Teradata) and make necessary adjustments.
Identify bottlenecks and implement performance improvements and optimizations in pipelines and large-scale data processing.
CI&T helps large companies transform AI potential into real business impact with AI deployment, AI-native execution, and tech-integrated business solutions. With 30 years of experience and 8,000 employees across 25+ countries, we accelerate innovation through expertise in agentic SDLC, application modernization, Data & AI, martech, and business strategy.
Build and optimize scalable data pipelines using Databricks, Spark, and PySpark for cloud migration.
Design and maintain data models (Bronze, Silver, Gold layers) to enable analytical insights.
Ensure data quality and governance while managing the technical roadmap and communicating with stakeholders.
CI&T helps large companies transform AI potential into real business impact with AI deployment and tech-integrated solutions. They have 8,000 employees across 25+ countries, fostering a collaborative and innovative culture.
Design, develop, and maintain scalable data pipelines and infrastructure.
Collaborate with cross-functional teams to ensure data integrity and usability.
Implement data quality and governance practices within data pipelines.
We help large enterprises transform with AI and tech-integrated solutions. With 8,000 employees across 25+ countries, we collaborate to build solutions with real impact.
Build and maintain data pipelines and infrastructure for digital analytics environments, ensuring robust ingestion, transformation, and availability of data.
Work closely with Martech experts to integrate data from sources like Google Analytics and Customer Data Platforms.
Manage platform data and technology, ensuring integrity, consistency, and alignment with business needs.
CI&T helps large enterprises harness AI for real business impact through AI deployment, AI-native execution, and tech-integrated business solutions. With 8,000 employees across over 25 countries, the company fosters a collaborative and innovation-driven culture focused on digital transformation.
Design, build, and maintain robust ETL/ELT processes for modern Data Lake architectures.
Write and optimize complex SQL queries using Python and PySpark for large-scale distributed data processing.
Collaborate with cross-functional teams to ensure data governance, quality, and observability across cloud-native platforms.
CI&T helps large enterprises transform the potential of AI into real business impact. With 30 years of experience, we are 8,000 CI&Ters across more than 25 countries.
Develop and evolve data ingestion applications using AWS tools such as Glue, Lambda, and SQS.
Optimize data pipelines with partitioning, skew balancing, and performance tuning.
Implement automated testing, code reviews, and deployments with a focus on FinOps.
CI&T helps large companies transform AI potential into real business impact through AI deployment, AI-native execution, and tech-integrated solutions. With 8,000 employees across 25+ countries, they foster a collaborative culture where AI is integrated into daily work.
Design and manage end-to-end data infrastructure including databases, Data Warehouses, Data Lakes, and Medallion Architecture.
Lead technical migration strategy from legacy systems to cloud platforms, ensuring smooth transition.
Translate business requirements into technical specifications, establish data models and governance policies.
CI&T helps large companies transform AI potential into real business impact with AI deployment, AI-native execution, and tech-integrated business solutions. With 30 years of experience and 8,000 employees across 25 countries, they collaborate to build solutions with real impact.
Design, implement, and maintain end-to-end data pipelines (ELT/ETL) with focus on reliability, reprocessing, and cost-efficiency.
Orchestrate data loads using Airflow (AWS) and serverless functions, model data in Snowflake (bronze/silver/gold) and develop transformations with dbt.
Build integrations and services in Python with FastAPI, write high-performance SQL, and ensure data quality, observability, security, and documentation.
CI&T helps large companies transform AI potential into real business impact with AI deployment, AI-native execution, and tech-integrated business solutions. With 30 years of experience and 8,000 employees across over 25 countries, they collaborate to build solutions with real impact.
Design and deliver complex data platform components or migration solutions on AWS using Databricks.
Troubleshoot performance, scalability, and reliability issues in cloud-native data environments.
Communicate technical topics clearly to stakeholders and produce high-quality documentation.
Caylent is an AI-first cloud services company that helps organizations turn ambitious ideas into meaningful business impact. As a fully remote global company with employees in Canada, the United States, and Latin America, they celebrate diverse cultures and foster a community of technological curiosity.
Design, optimize, and maintain robust and scalable data pipelines on AWS cloud.
Plan and execute migration of systems and data volumes to new cloud infrastructure.
Collaborate with multidisciplinary teams to implement efficient data solutions and ensure data governance.
CI&T helps large companies transform AI potential into real business impact with AI deployment, AI-native execution, and tech-integrated business solutions. With 30 years of experience and 8,000 employees across 25+ countries, they collaborate to build solutions with real impact.
Build and operate reliable data infrastructure supporting marketing, payments, and product analytics.
Design and maintain data pipelines using Databricks and Apache Airflow to ensure accurate and timely reporting.
Collaborate with cross-functional teams to translate business needs into scalable data solutions.
Our partner company builds digital products for the mobile app industry, focusing on data-driven marketing and analytics. They offer a collaborative, international environment with a remote-first culture and emphasize continuous learning and professional growth.
Lead design and implementation of enterprise data pipelines within Microsoft Fabric.
Mentor data engineers and establish best practices for data engineering.
Ensure data platform security, availability, and performance requirements are met.
Rackner is a software consultancy focused on building mission-critical systems for the U.S. government. The company provides a creative and forward-thinking environment with growth opportunities and modern perks.
Design and implement the UniForm write layer (Delta + Iceberg dual metadata).
Build GCS → BigQuery ingestion pipelines for structured operational datasets.
Develop and implement in Python and Spark, including Kafka-based CDC patterns for real-time ingestion.
Railroad19, Inc develops customized software solutions and provides software development services. They are a specialized team of developers and architects with a culture of hard work and industry leadership, valuing employees and offering remote work from the US.
Design and maintain large-scale cloud data infrastructure for healthcare applications.
Build efficient ETL pipelines, self-service tools, and microservices using Azure, Snowflake, and Databricks.
Collaborate with product owners and team leads to define efficient data pipelines and schemas.
Sigma Software is an IT solutions company delivering innovative technology to global clients across multiple industries. They foster a supportive environment with cutting-edge technologies and meaningful projects.
Design and drive the execution of a scalable, cloud-native target-state architecture for the data platform.
Define architectural patterns for scalable automated data ingestion and deliver hands-on proof-of-concepts using SQL, PySpark, and Serverless Python.
Drive governance, security, and engineering excellence across all platform components, ensuring alignment with business and product strategies.
G-P is a SaaS-based Global Employment Platform that enables clients to expand into over 180 countries. Our diverse, remote-first teams foster innovation and value every contribution, empowering employees with flexibility and resources.
Design, build, and maintain scalable data pipelines and ETL/ELT processes using Databricks and Spark.
Architect and optimize data models and storage solutions for analytics and operational use.
Implement observability, alerting, and data quality monitoring for critical pipelines.
Sonatype provides end-to-end software supply chain security solutions, protecting against malicious open source and managing SBOMs. With over 2,000 organizations and 15 million developers using its platform, it focuses on innovation and security in software development.
Develop and maintain data pipelines using Databricks, PySpark, Python, and SQL to transform raw financial data into reliable, curated datasets.
Integrate new data sources, support financial reconciliation, P&L routines, and global closing processes, ensuring data consistency and accuracy.
Collaborate with multidisciplinary teams to deliver high-quality data solutions, support AI initiatives, and implement continuous improvements in data performance and quality.
CI&T helps large companies transform AI potential into real business impact with AI deployment, AI-native execution, and tech-integrated business solutions. With 30 years of experience and 8,000 employees across 25+ countries, they collaborate to build solutions with real impact.
You will work under the guidance of the Central Platform Team to build our Databricks engineering framework.
You will establish reusable development patterns and advance our core codebase.
You will use modern data engineering practices and AI-assisted tools to build scalable data solutions.
Talan is an international consulting group specializing in innovation and business transformation through technology. With over 7,200 consultants in 21 countries and a turnover of €850M, they are committed to delivering impactful, future-ready solutions.
Develop and maintain data pipelines for data ingestion, transformation, and processing using Python and SQL.
Collaborate with Data Engineers, Analysts, and global stakeholders to deliver effective data solutions.
Apply modern data engineering practices, including Git, CI/CD, and cloud platforms like Azure and Databricks.
They partner with global clients to deliver large-scale digital transformation initiatives through data solutions. They foster a remote, multicultural environment emphasizing technical growth, collaboration, and continuous improvement.