The Role:
- Design, build, and optimize modern cloud-based data platforms powering analytics, AI, and data products.
- Work on scalable batch, streaming, and near-real-time pipelines, with a focus on data governance and observability.
- Collaborate with data scientists and ML engineers to support GenAI and machine learning workflows.
Responsibilities:
- Implement lakehouse architectures and medallion layering in Databricks or Fabric.
- Orchestrate workflows using Airflow, dbt, or native tooling, ensuring automation and reliability.
- Enable data serving layers and ensure data quality, lineage, and access control.
Must-Have Qualifications:
- Strong hands-on experience with Apache Spark, Delta Lake, Python, and SQL.
- Deep expertise in either AWS data services with Databricks or Azure/Fabric with Databricks.
- Bring curiosity to business problems, applying analysis and recommendations to shape requirements.
Benefits:
- Private health insurance, education program, and wellbeing program.
- 24 vacation days, free beverages, and company events.
- Challenging projects, cool colleagues, and honest feedback culture.