Source Job

$112,000–$179,000/yr
US

  • Build out and configure a dedicated Government Databricks workspace within the existing environment
  • Design and implement data ingestion pipelines from core agency systems including financial, HR, CRM, and ITSM
  • Implement Databricks Unity Catalog for centralized data governance, metadata management, and end-to-end lineage tracking

PySpark Spark SQL Python Unity Catalog Delta Lake

20 jobs similar to Databricks Data Engineer

Jobs ranked by similarity.

$140,000–$210,000/yr
US

  • Design and build ingestion pipelines from enterprise source systems into the Databricks lakehouse using Delta Lake, owning Bronze-layer ingestion and building Silver-layer pipelines for cleansing and standardization.
  • Implement and maintain Databricks platform constructs for secure delivery, including catalogs, schemas, service principals, and job orchestration, and build CI/CD pipelines for data platform assets.
  • Apply data classification and segregation requirements within pipeline design, build data quality controls reflecting business meaning, and partner with teams to ensure reliable, governed data for downstream use.

Shield AI is a venture-backed defense-tech company founded in 2015 with the mission of protecting service members and civilians with intelligent systems, developing products like Hivemind autonomy software and V-BAT and X-BAT aircraft. The company has offices and facilities across the U.S., Europe, the Middle East, and Asia-Pacific, and its technology actively supports operations worldwide.

  • Design and implement robust, scalable data ingestion and transformation pipelines using Databricks, PySpark, and distributed processing.
  • Implement Delta Lake principles focusing on CDC and schema evolution, and integrate data quality frameworks within CI/CD pipelines.
  • Develop and optimize complex SQL and Python scripts, handling diverse data sources and supporting data governance solutions.

Mobile Wave Solutions is a professional services company specializing in software development as a service. With a team of over 120 engineers, we deliver scalable, high-quality software that empowers our global clients to innovate and grow.

$101,144–$174,591/yr
US

  • Lead data engineering efforts within Palantir Foundry, including ontology design, pipeline development, and data integration for Army AI2C mission applications.
  • Build and maintain data engineering workflows in Databricks, including notebooks, Delta Lake tables, and Spark jobs.
  • Provide technical leadership and mentorship to grow team capabilities across Foundry and Databricks.

LMI is a digital solutions provider dedicated to accelerating government impact with innovation and speed. The company serves defense, space, healthcare, and energy sectors, and is headquartered in Tysons, Virginia.

UK

  • Design and maintain scalable batch and near real-time data pipelines in Databricks, integrating data from various sources.
  • Build unified customer profiles and support identity resolution, deduplication, enrichment, and Golden Record creation.
  • Enable CRM and marketing teams to access trusted, activation-ready customer data with minimal latency.

Massive Rocket is a rapidly scaling Braze and Snowflake agency that transforms how digital marketing, product, and engineering teams connect. They have grown quickly in five years and are aiming for $100M in revenue, with a culture of ownership, collaboration, and growth.

US Canada

  • Build and maintain batch data pipelines ingesting from Workday and legacy systems into a medallion lakehouse on Azure Databricks and Microsoft Fabric.
  • Develop transformation logic across Bronze/Silver/Gold layers using Spark, Python, and SQL, and implement data-quality checks.
  • Work within a federated architecture and Unity Catalog security model, contributing to documentation and code reviews.

CampusWorks is a consulting firm that helps higher education institutions leverage technology and managed services to drive transformation and student success. Founded in 1999, it is a large, virtual company with a culture that values work-life balance and employee appreciation.

$130,000–$165,000/yr
US

  • Design, build, and maintain scalable data pipelines and systems using Python, PySpark, and AWS-native services.
  • Develop and optimize data platforms on AWS, leveraging technologies like Redshift, Glue, and Athena.
  • Mentor junior engineers and collaborate with cross-functional teams in an Agile environment.

Capital Technology Group provides expert consulting services in software development, digital transformation, and data analytics. Recognized as a Top Workplace by The Washington Post, they foster a culture where employees feel trusted and empowered.

$0–$215,000/yr
US Unlimited PTO

  • Lead, coach, and develop a global team of data engineers while guiding technical architecture and delivery.
  • Build and scale large-scale data pipelines, production datasets, and QA systems using Databricks, Airflow, SQL, and PySpark.
  • Partner with product and research teams to transform complex data into reliable, production-grade assets for customers and internal applications.

YipitData is a leading market research and analytics firm that analyzes billions of alternative data points to provide actionable insights for the disruptive economy. The company recently raised $475M, operates globally with offices in the US, APAC, and India, and has been recognized by Inc. as a Best Workplace for three consecutive years.

US Canada

  • Lead end-to-end data governance workstreams including data cataloging, data quality, metadata management, lineage, and access control frameworks.
  • Translate client business needs into governance solutions and roadmaps without requiring significant oversight.
  • Create, socialize, and implement data governance policies, procedures, and frameworks for clients while ensuring compliance.

Lovelytics helps enterprises modernize their data and implement AI, delivering over $2B in business value for Fortune 500 clients. They have grown from 85 to 500+ people across the Americas, maintaining a technically excellent, low-ego culture that earned them a Best Places to Work recognition.

  • Design and deploy highly performant end-to-end data architectures using Databricks and Apache Spark.
  • Develop and manage scalable batch and streaming solutions in cloud ecosystems like AWS, Azure, or GCP.
  • Lead large-scale data engineering projects, ensuring technical delivery and client management.

Derex Technologies Inc provides IT consulting, staffing solutions, and software services to global clients. With over two decades of experience, the company delivers high-quality technology professionals across various industries.

Global

  • Design, develop, and maintain data pipelines using Azure Databricks.
  • Build and optimize data transformations using PySpark and SQL in Databricks.
  • Implement and maintain Lakehouse architectures using Delta Lake.

Miratech helps visionaries change the world. We are a global IT services and consulting company with nearly 1000 full-time professionals and a culture of Relentless Performance.

Slovakia

  • Own production data pipelines end-to-end, ensuring reliability, performance, and scalability. - Build and optimize data solutions using PySpark and Foundry tools to process massive datasets. - Champion data quality by designing monitoring and validation frameworks to guarantee trusted data.

We are an IT solutions company providing innovative information and communication technology services. We have grown to over 3,900 employees and are the second largest employer in eastern Slovakia, with a culture focused on continuous improvement and transformation.

$170,000–$230,000/yr
US Unlimited PTO

  • Build, expand, and optimize data infrastructure to create the most accurate dataset of identities and their relationships.
  • Develop and operate secure, scalable, and reliable data ingestion and ETL/ELT pipelines that meet product requirements.
  • Design and maintain a data observability framework to ensure data meets strict quality and freshness standards.

SentiLink provides innovative identity and risk solutions, empowering institutions and individuals to transact with confidence. The company is growing rapidly, has verified hundreds of millions of identities, and is backed by top investors like Craft Ventures and Andreessen Horowitz, with offices across the US and India.

$85,000–$141,000/yr
US

  • Design and optimize data pipelines and applications to support operational decision-making.
  • Apply industrial engineering and operations research techniques to improve efficiency and readiness.
  • Collaborate with stakeholders to deliver data-driven insights and support digital transformation.

Brazil

  • Own and evolve notebooks across the full medallion architecture, from raw ingestion through a fully modeled dimensional layer.
  • Design and implement dimensional modeling artifacts such as facts, dimensions, and slowly changing dimensions for downstream business intelligence.
  • Serve as the primary technical reference, reviewing pull requests, mentoring team members, and proposing architectural improvements.

CI&T helps large enterprises transform the potential of AI into real business impact with AI Deployment, AI-native execution, and tech-integrated business solutions. With 30 years of experience and 8,000 employees across 25 countries, we foster a collaborative culture that accelerates innovation in Agentic SDLC, Application modernization, Data & AI, Martech, and Business strategy.

US

  • Administer and enhance the Advana data platform for enterprise cybersecurity reporting and analytics.
  • Develop and maintain secure data ingestion pipelines and executive dashboards using Databricks and Qlik Sense.
  • Ensure data governance and compliance with DoD standards while supporting operational transitions and knowledge transfer.

Spry Methods supports the United States Marine Corps by modernizing mission-critical cybersecurity analytics dashboards. The team focuses on building secure, scalable solutions for enterprise data platforms and reporting, operating within a collaborative and mission-driven culture.

UK Ireland

  • Design, build, and operate scalable data pipelines using modern ETL/ELT approaches on cloud-native platforms.
  • Work with technologies such as Databricks, PostgreSQL, and distributed processing frameworks to deliver production-ready solutions.
  • Collaborate with analysts, data scientists, and AI engineers while mentoring junior team members.

Version 1 is a technology and transformation solutions company trusted by global brands. With over 3,300 employees and €350m revenue, they are a values-driven employer focused on employee wellbeing and professional growth.

LATAM

  • Support Snowflake-to-Databricks data migration activities.
  • Design and implement scalable data pipelines using Snowflake, Databricks, SQL, Python, and dbt.
  • Optimize data workflows for performance, scalability, cost efficiency, and maintainability.

Hiflylabs is a Budapest-based company delivering Data Warehouses, Business Intelligence, and Data Analytics solutions. With over 250 employees and 10 years of experience, they serve financial, telecommunication, and energy sectors, fostering a culture of innovation and collaboration.

Brazil

  • Oversee the full lifecycle of data products within the Data Core team, from ideation to operation and continuous optimization.
  • Ensure the data infrastructure (Databricks, Unity Catalog, data lake) operates with excellence, reliability, and scalability.
  • Lead and develop a cross-functional team of data engineers, scientists, and analytics experts, balancing long-term vision with hands-on execution.

We help large companies transform AI potential into real business impact with AI deployment, AI-native execution, and tech-integrated business solutions. With 30 years of experience, we are 8,000 CI&Ters in over 25 countries, collaborating to build solutions with real impact.

$108,400–$180,700/yr
US 4w PTO

  • Design and build production data pipelines on Databricks using medallion architecture and Data Vault 2.0.
  • Harden curated and presentation layers to drive donor lifetime value, retention, and campaign ROI.
  • Build identity resolution and data observability to ensure trustworthy data for 24,000+ nonprofits and AI agents.

Bloomerang offers a giving platform and support for nonprofits to raise more and retain donors. They serve tens of thousands of nonprofits and have a culture built on core values of Simplify, Care, and Act.

Ukraine

  • Design and build batch data pipelines that ingest, validate, and transform multi-billion-row datasets.
  • Model complex real-world data including dimensional models and temporal data.
  • Develop and operate workloads on lakehouse platforms like Databricks, Spark, and Delta.

Simulmedia builds an advanced TV and streaming advertising platform. They have a team of engineers, data scientists, and designers who are obsessed with building cutting-edge technology.