Data Engineers
22199
7 Sep, 2026 to 31 Aug, 2029
We are looking for 4 Data Engineer to migrate SAS workloads to the Databricks Lakehouse Platform. You will translate legacy SAS data processing into PySpark and SparkR, design scalable pipelines, and optimize Delta Lake tables for performance and reliability. Collaborate with analysts and data scientists to ensure functional equivalence between SAS and Databricks implementations, implement CI/CD with Git and Azure DevOps, and validate data through automated testing and data quality checks.
Key Responsibilities
Migrate SAS workloads to the Databricks Lakehouse Platform.
Translate legacy SAS data processing into PySpark and SparkR
Design scalable pipelines
Optimize Delta Lake tables for performance and reliability
Collaborate with analysts and data scientists to ensure functional equivalence between SAS and Databricks implementations
Implement CI/CD with Git and Azure DevOps, and validate data through automated testing and data quality checks.
Required Experience Mandatory:
Databricks
Good experience to have:
PySpark
SAS, Basics
ETL
SQL
Language requirements:
Finnish mandotary
Start date: As soon as possible after candidate is choosen
End date: 31.12.29 strong possibility for extension