Senior Data Engineer
Core
Design, build, and maintain scalable data pipelines and lakehouse structures to support analytics, BI, machine learning, and Generative AI applications.
Role type
Senior Data Engineer (Generative AI & Lakehouse)
Builds
Production-ready data solutions, curated knowledge datasets, and data pipelines for GenAI use cases (RAG, agent workflows).
Domain
Enterprise Data Engineering, Generative AI, Lakehouse Architecture
Deliverable
production ML models | product features | infrastructure
Required skills
Python, SQL, Apache Spark, Databricks (Delta Lake, Unity Catalog), Data Lakehouse Architecture, CI/CD, Automated Testing, Data Governance, Incident Resolution, Root Cause Analysis
Preferred skills
Generative AI workloads, RAG data patterns, Vector/embedding workflows, Cloud-native platforms (AWS/Azure), Analytics operational workloads
Technologies
Databricks, Delta Lake, Unity Catalog, Apache Spark, Python, SQL
Responsibilities
Design and maintain scalable data pipelines and lakehouse structures; Build data pipelines for Generative AI applications including curated datasets and lineage management; Enable GenAI data patterns including RAG and prompt/context preparation; Develop production-grade pipelines using Python, PySpark, and SQL; Implement automated testing and CI/CD practices; Support operational stability, incident resolution, and root cause analysis.
Seniority
Senior, hands-on IC