Similar Jobs
See allData Engineer (Remote Opportunity)
VetsEZ
US
Python
SQL
Azure
Senior Data Engineer
540
US
Python
SQL
Apache Spark
Data Engineer
CampusWorks
US
Apache Spark
Python
SQL
Databricks Data Engineer
Peraton
US
PySpark
Spark SQL
Python
Junior Data Engineer (Azure Databricks)
Miratech
Global
Azure Databricks
PySpark
Spark SQL
Primary Responsibilities:
- Design, build, test, and maintain scalable data pipelines and ETL/ELT processes for cloud-based data environments.
- Implement patient-matching and record-linkage logic across multiple identifiers for accurate Veteran data.
- Harmonize datasets across VA facilities, resolving structural and format inconsistencies.
Minimum Qualifications:
- 8+ years of relevant experience in data engineering with ETL/ELT pipelines and cloud data environments.
- BA/BS degree in Computer Science, Information Systems, or related technical field.
- Prior experience supporting federal agency programs; must be a U.S. citizen or authorized to work in the U.S.
Desired Qualifications:
- Hands-on experience with Azure Databricks and PySpark for large-scale distributed data processing.
- Proficiency in Python for pipeline automation and data wrangling within federal health data contexts.
- Experience developing Power BI reports and dashboards using Databricks SparkSQL.
Aptive
Aptive partners with federal agencies to achieve their missions through improved performance, streamlined operations and enhanced service delivery. Founded in 2012, the company employs over 300 people nationwide and focuses on human-centered services.