Design and develop robust data pipelines using Databricks and AWS S3 while implementing the Lakehouse Medallion Architecture. Integrate and centralize data from multiple internal systems while performing independent unit testing and data validation.
This is a remote position.
We are looking for a Senior Data Engineer to design and develop our lakehouse architecture and data pipelines as the organization transitions from its current Azure DevOps environment to Databricks on AWS S3.
The role will focus on building and maintaining a scalable data platform based on the Databricks Lakehouse Medallion Architecture, supporting the centralization and integration of data across our in-house systems.
This Role requires overlap with US Business hours
Key Responsibilities
Design and develop robust data pipelines using Databricks and AWS S3.
Contribute to the implementation and evolution of the Lakehouse Medallion Architecture.
Integrate and centralize data from multiple internal systems.
Understand and work with ERDs and data models, including concepts such as data grain and join cardinality.
Develop, validate, and troubleshoot data transformations and pipelines.
Perform unit testing and data validation independently after development.
Collaborate closely with teams working across the US time zone.
Requirements
Strong experience as a Senior Data Engineer.
Hands-on experience with Databricks, AWS S3, and modern data lakehouse architectures.
Solid understanding of data modeling, ERDs, data grain, and join cardinality.
Experience designing and developing production-grade data pipelines.
Ability to independently perform unit tests and validate developed solutions.
English at C1 level, with strong communication skills.
Ability to work remotely from the EU while maintaining US working hours / significant US time-zone overlap.
“I was the first applicant for a remote marketing position that got listed on the company website the same day I applied. Had an interview within 48 hours!”