IT engineer data lakehouse
Core
Design, develop, and operate scalable data pipelines and data products in Azure Databricks to enable data-driven decision-making for HR, Purchasing, and Finance domains.
Role type
Senior Data Engineer (Data Lakehouse)
Builds
Scalable batch and streaming pipelines, bronze/silver/gold data layers, curated data marts, and dimensional models for enterprise reporting.
Domain
Automotive industry data engineering and analytics
Deliverable
production ML models | product features | dashboards & analysis
Required skills
PySpark, Scala, Azure Databricks, CI/CD automation, test-driven development, data modeling (star/snowflake, 3NF), master data management, performance tuning, version control
Preferred skills
mentoring junior developers, leading implementation workstreams
Technologies
Azure Databricks, PySpark, Scala, SAP, Power BI, Unity Catalog
Responsibilities
Design scalable batch and streaming pipelines; implement ingestion from structured and semi-structured sources; build lakehouse layering architecture; implement dimensional models; develop automated data validation tests; monitor pipeline health and optimize cluster resource usage; refactor legacy pipelines; document pipeline logic and technical decisions; align designs with governance and metadata standards.
Seniority
Senior, hands-on IC